Blog

Guides, tutorials, and insights on AI coding tools and API providers.

· 6 min read

DeepSeek R1: OpenAI-Level AI, Now Open Source

Oh, speaking of last night—I stayed up until 3 a.m. Guess what I was doing? Not working overtime, not binge-watching shows, not gaming. I was "playing" with a m

Read more →
· 6 min read

Have you ever seen AI drawing crash and burn

Have you ever seen AI drawing crash and burn? I have. More than once. Back in late 2022, when Stable Diffusion first blew up, I spent every day tweaking prompts

Read more →
· 6 min read

A 160-Line Prompt That Suddenly Made GPT-4o and Claude 3.5 Click

Okay, got the instructions. As an editor, I'll first verify the facts, then polish the text, remove the AI tone, and make the article sound like a real person t

Read more →
· 6 min read

Claude Code Source Code: A Deep Dive for Devs

You see, my attitude toward Claude Code went through three complete faceplants. The first time was when it first blew up. I glanced at it and thought: Isn't thi

Read more →
· 6 min read

First Gate: What exactly does model parallelism parallelize? You're probably half wrong!

Let me first tell you a story that gave me the "creeps." What was your first reaction when you saw Megatron's model parallelism code? Here’s mine: At that point

Read more →
· 6 min read

Just half a year ago, I was absolutely fuming at a model that could only chat

You know what? Just half a year ago, I was absolutely fuming at a model that could only chat. I said to it: "Can you help me organize this week's code commit hi

Read more →
· 6 min read

I'll Be Straight with You: What Really Changed in RAG 2.0? Stop Letting Those "RAG Is Dead" Headlines Fool You!

Let me tell you something embarrassing. At the end of 2024, I confidently wrote several "RAG technology summaries" with absolute certainty—RAG had peaked! Hit t

Read more →
· 6 min read

U-Net Converges Twice as Fast as Transformers in Diffusion Models

Did you know? Last year, when I was trying to get an SD inpainting project to work, I almost smashed my computer! There are so many tweaked versions of U-Net ou

Read more →
· 4 min read

Hey, I know what you're about to say

Hey, I know what you're about to say! MCP is taking the world by storm, right?! Tutorials, guides, from beginner to expert—everywhere you look... Stop! Hold on

Read more →
· 7 min read

Alignment and Fusion Are Not the Same Thing

One late night last summer, I ran an experiment. I fed a video clip to three models: one looked at the visuals, one listened to the audio, and one read the subt

Read more →
· 4 min read

Tsinghua Post-00s Uncover the Culprit Behind AI Hallucinations: LLMs Aren’t Stupid—They’re Just Tryi

Translate to English, keep the storytelling style: Oh my god, you might not believe this— The other day, I came across a Tsinghua research paper and almost jump

Read more →
· 6 min read

1. Three Parallelism Strategies, Plainly Three Cuts

It took me three months to finally understand what distributed training of large models is all about. Let me tell you a story first. A friend of mine, full of a

Read more →
← Previous 1 ... 24 25 26 27 28 ... 42 Next →