Blog

Guides, tutorials, and insights on AI coding tools and API providers.

· 6 min read

Large Model Inference Optimization: The Pits I Fell Into, and You Might Be Jumping Right In

To be honest, I held off writing this article for over half a year before I finally dared to put pen to paper. The reason is simple—too many people oversimplify

Read more →
· 5 min read

Brothers, fine-tuning large models? I'm totally sold on it!

After checking, the original facts are basically accurate, the data calculations are reasonable, and there are no major errors that need correction. AI-style ex

Read more →
· 6 min read

Oh my god, when it comes to the pacing of web novel writing, I have so many bitter tears to shed!

Let me start by telling you a story—my story. I wrote columns for ten years. The first three years were such a disaster that even my own mother would have wante

Read more →
· 4 min read

Policy Gradient, PPO, and PPG: A Complete Guide

Five years ago, I encountered reinforcement learning for the first time. Not in a school lab, not while grinding LeetCode—it was on a robotic arm grasping proje

Read more →
· 5 min read

LLM Inference Quantization: FP8 vs INT8

Let me start with a true story. Two years ago, I was working on inference services, back when the A100 was still the hot commodity. I was messing around with Ll

Read more →
· 5 min read

LLM Training Made Easy: From GPU Memory to DeepSpeed

Alright, leave it to me! I'm going to turn this article inside out and breathe a brand new soul into it. Last month, I was hunkered down in the server room, sta

Read more →
· 6 min read

Deconstructing Where ChatGPT's Abilities Come From: A Veteran's Hands-On Notes

I've been writing this column for ten years, and I've seen plenty of tech articles that talk big. But this ChatGPT thing... guess what? It's genuinely different

Read more →
· 6 min read

Final Version

Okay, no problem. As an editor, I'll carefully fact-check, correct data, and refine the language style to make the article more solid and the rhythm more natura

Read more →
· 3 min read

LLaMA, ChatGLM, BLOOM: Top Open-Source LLMs Compared

Last year, I was full of confidence—I got my hands on the original LLaMA-7B and wanted to play around with Chinese instructions. And guess what? The sentences t

Read more →
· 6 min read

Stop Guessing, Just Teach It! A Discovery That Almost Made Me Give Up on Prompt Tuning

Have you ever tried shouting at something for ages, only to have it completely fail to understand what you're saying? Last year, I was working on a zero-shot se

Read more →
· 6 min read

The Pitfalls of Distributed Training: It Took Me Two Whole Weeks to Finally Understand TP, SP, and Communication Overlap

Hey, friend! Let me tell you about a topic I've got a love-hate relationship with—distributed training. Two years ago, when I first wrote Tensor Parallel code

Read more →
· 5 min read

Section 1: Is Emergence Actually Pseudoscience?

Alright, today let's talk about a seriously thrilling topic—emergence. You know, if you scroll through your phone, you've definitely seen the word: "AI emerged

Read more →
← Previous 1 ... 27 28 29 30 31 ... 42 Next →