Working Notes on Agent Systems/Brad Zhang

Topic desk

Memory

What deserves to survive between sessions, and what should be compacted, rewritten, or deliberately forgotten?

Memory is treated here as an editorial surface rather than a warehouse. The question is less about infinite recall than about what changes the next judgment.

Live source unreachable; showing the most recent snapshot.

Longform

All longform

OpenAI Proactively Discloses Long Range Agent Tasks Out of Bounds: The Safe Steering Behind PR # 287

A single tool approval begins to lose global semantics when the model can be tried for hours on end. On July 20, 2026, OpenAI unveiled an on-premises pause. A generic model that...

Jul 21, 2026 · 12 min read

Everyone is interested in the fact that GPT 5.6 sol output speed is 12 times faster than Kimi K3.

Everyone is interested in the fact that GPT 5.6 sol output speed is 12 times faster than Kimi K3. Let me list the relevant materials to read together: Group 1: Where...

Jul 20, 2026 · 2 min read

How to work with GPT 5.6 Sol together with Kimi K3? This is my suggestion after testing...

How to work with GPT 5.6 Sol together with Kimi K3? This is my suggestion after testing. In a word: Sol is a better "brain", K3 is a cheaper "stamina + long memory". The matchin...

Jul 20, 2026 · 3 min read

The output speed of GPT 5.6 sol is approximately 12 times that of Kimi K3! Everyone feels that K3 is much

The output speed of GPT 5.6 sol is approximately 12 times that of Kimi K3! Everyone feels that K3 is much slower than GPT 5.6 sol. This is indeed the case in my experience...

Jul 20, 2026 · 2 min read

A summary of the best articles on Graph Engineering! Prompt, the next engineering parad...

A summary of the best articles on Graph Engineering! Prompt, the next engineering paradigm from Loop to Graph, determines how the Agent starts. Context determines what the Agent...

Jul 20, 2026 · 4 min read

Why does Codex crash every time I open it? Finally found the reason! Because I am a heavy user of Codex,

Why does Codex crash every time I open it? Finally found the reason! Because I am a heavy user of Codex, I usually start with 6G of memory, and sometimes it will...

Jul 19, 2026 · 2 min read

ChatGPT Work: AI starts delivering work for you In the past, we used ChatGPT to get answers.

ChatGPT Work: AI starts delivering work for you In the past, we used ChatGPT to get answers. Now, OpenAI is thinking about another thing: letting AI pick up a complete target, f...

Jul 19, 2026 · 1 min read

The Agent's memory should not be stuffed into the context window

When many people discuss coding agents such as Codex and Claude Code, their first reaction is model capabilities: whether they can read code, correct bugs, and...

Jul 17, 2026 · 8 min read

I connected Codex and Cursor: an automatic memory bus allows the two Agent software to communicate freely

Recently I am doing a very specific thing: using the source code of pi agent and a set of small exercises. I hope that Codex will be responsible for the reading history, combing...

Jul 14, 2026 · 11 min read

Personally tested it and it works! From now on, you have a smart model routing locally, which is fast and

Personally tested it and it works! From now on, you have a smart model routing locally, which is fast and saves money. 💰 The model routing is working well this time. It...

Jul 14, 2026 · 3 min read

So the question is, who will be the third child of the Yusan family? 🤔 On Claude's side, Sonnet 5, Fable 5

So the question is, who will be the third child of the Yusan family? 🤔 On Claude's side, Sonnet 5, Fable 5, and Mythos 5 have already rolled out the model echelon, and...

Jul 10, 2026 · 4 min read

gpt 5.6 sol is loaded, reset, start biu biu biu Don't go to the typhoon on the weekend,...

gpt 5.6 sol is loaded, reset, start biu biu biu Don't go to the typhoon on the weekend, just stay home and code with 5.6 sol. What is the biggest difference between gpt 5.6 sol...

Jul 9, 2026 · 2 min read

Anthropic’s 85-minute Fable 5 workshop explains how to implement the next generation AI Agent

Video entrance Opening keynote: https://www.youtube.com/watch?v=GMIWm5y90xA The capability curve: https://www.youtube.com/watch?v=tP4MGcJ80Y0 How to get to production...

Jul 8, 2026 · 18 min read

The significance of J-space is to give the previously black-box model reasoning an opportunity for

The significance of J-space is to give the previously black-box model reasoning an opportunity for observability and human intervention and control, so that corrections can...

Jul 7, 2026 · 2 min read

The term software factory has finally begun to move from a slogan to engineering sites.

Original video: WF2026: Software Factories & Keynotes ft. Microsoft, OpenAI, OpenClaw, Z.ai (GLM), MiniMax, HF Link: https://www.youtube.com/watch?v=htM02KMNZnk Start I just fin...

Jul 4, 2026 · 8 min read

After pre-training, what does the next training paradigm for AI look like?

Source Original video: What does the next training paradigm look like? Channel: Dwarkesh Patel Link: https://www.youtube.com/watch?v=20p5-kQXF_Q Start In this video, Dwarkesh Patel

Jul 2, 2026 · 15 min read

From Prompt Engineering to Harness Engineering: four evolutions of the AI ​​engineering system

The rapid development of products such as Agent, Codex, and Claude Code has led to the emergence of a number of new terms in the field of AI engineering: Prompt...

Jun 19, 2026 · 7 min read

Summary of the best articles of Loop Engineering! After Agent began to focus on long tasks in 2026,

Summary of the best articles of Loop Engineering! After Agent began to focus on long tasks in 2026, the focus slowly became: How to design a loop system that can...

Jun 18, 2026 · 2 min read

If you want to truly understand how LLM is trained, CS336 is the most worthy set of open courses to learn the system.

If you want to truly understand how LLM is trained, CS336 is the most worthy set of open courses to learn the system. Building a modern LLM engineering stack from scratch covers...

Jun 18, 2026 · 2 min read

Let's say you're the head of technology at a technology company. Last year, you used GraphRAG to build an

Let's say you're the head of technology at a technology company. Last year, you used GraphRAG to build an internal knowledge base and used 20,000 contracts to do PoC. The...

Jun 17, 2026 · 17 min read

When dynamic workflow and front-end design are combined, there will be unexpected effects!

When dynamic workflow and front-end design are combined, there will be unexpected effects! Share an open source project: fireworks-design This project is easily misunderstood as...

Jun 16, 2026 · 5 min read

How to burn token series? (1) Use Dynamic Workflow to let 100+ agents work in parallel

If you look at the picture below, it should basically cover the daily usage scenarios of ClaudeCode or Codex partners. These are some intermediate processes of a website...

Jun 15, 2026 · 4 min read

Claude Fable is for enterprises to replace traditional mid-to-high-level development, and the ROI is positive.

Claude Fable is for enterprises to replace traditional mid-to-high-level development, and the ROI is positive. After the release of the Claude Fable model, I saw many crazy exam...

Jun 10, 2026 · 4 min read

Agent Loop is popular🔥: Agent goes from Demo to production with a Loop in between

There has been a heated debate in overseas AI circles these days. Matt Van Horn sorted out this debate in “WTF Is a Loop? Peter Steinberger vs. Boris Cherny”: You shouldn’t be p...

Jun 9, 2026 · 10 min read

Five Traps of Agent Harness 1. Self-evaluation is a trap. Use an adversarial evaluator....

Five Traps of Agent Harness 1. Self-evaluation is a trap. Use an adversarial evaluator. Self-evaluation is a trap. Use an adversarial evaluator. Many teams will do this: Agent g...

Jun 8, 2026 · 3 min read

Agentic Workflow is the most common enterprise scenario. Today I read a sharing about h...

Agentic Workflow is the most common enterprise scenario. Today I read a sharing about how to define a stable and reliable Agentic Workflow within the enterprise. To summarize br...

Jun 5, 2026 · 2 min read

From Device Plugin to DRA, I systematically reviewed the evolution route of Kubernetes device resource management.

From Device Plugin to DRA, I systematically reviewed the evolution route of Kubernetes device resource management. Hardcore post from AI Infra! Collect it first and then...

Jun 3, 2026 · 3 min read

Pi Agent = customizable coding harness / agent workbench Pi Agent can be understood as a minimalist, self-transformable terminal coding agent harness.

Pi Agent = customizable coding harness / agent workbench Pi Agent can be understood as a minimalist, self-transformable terminal coding agent harness. It is not a general orches...

Jun 2, 2026 · 2 min read

Asynchronous Agent Era: From Code Assistant to Software Factory

I have reorganized the content of this podcast of Latent Space. It is worth taking a look at the in-depth discussion and sharing related to asynchronous Agent. I also...

May 31, 2026 · 22 min read

From the evolution of synchronous Agent to asynchronous Agent, I systematically sorted out the relevant content!

From the evolution of synchronous Agent to asynchronous Agent, I systematically sorted out the relevant content! The watershed between Async Agent and ordinary Agent is not "whe...

May 31, 2026 · 4 min read

This is one of the reasons why I love Codex 🔥 Many old users have regarded "Codex reset" as a new holiday in

This is one of the reasons why I love Codex 🔥 Many old users have regarded "Codex reset" as a new holiday in the AI ​​engineering circle 😂 Especially in high-intensity...

May 27, 2026 · 2 min read

My exploration and practice sharing in the past two months: Minimum engineering closed loop for long task Agent

A picture shared by engineer Claude accurately describes the three most common problems of long-task agents: the inability to keep track of the state, the inability...

May 25, 2026 · 27 min read

Ten times better, not ten times harder - the logic of growth in the AI ​​era

There is a question that has always made me feel a little strange: many people regard AI as a speed-increasing tool. It used to take half an hour to write an email, now...

May 23, 2026 · 7 min read

The main line of Agent development in 2026: moving from Agent Framework to Agent Runtime.

The main line of Agent development in 2026: moving from Agent Framework to Agent Runtime. The most obvious change in 2026 is that the Agent technology stack begins to be layered...

May 22, 2026 · 2 min read

FSD enters China: Autonomous driving begins to fight the "long-tail war". After China is included in Tesla's

FSD enters China: Autonomous driving begins to fight the "long-tail war". After China is included in Tesla's FSD Supervised country list, the focus of autonomous driving...

May 21, 2026 · 5 min read

Visualize that with the rapid development of agent cli and harness, the interaction and logic core will soon

Visualize that with the rapid development of agent cli and harness, the interaction and logic core will soon be reconstructed by Golang, which is very suitable for...

May 19, 2026 · 1 min read

Don't let the AI Agent step on the same pit a second time. I recently saw an interestin...

Don't let the AI Agent step on the same pit a second time. I recently saw an interesting open source project: RoBrain RoBrain is the "decision memory layer" of the AI programmin...

May 19, 2026 · 2 min read

How should you combine goal and display using subagent for codex? along with their technical principles and

How should you combine goal and display using subagent for codex? along with their technical principles and some typical scenarios. goal is responsible for “what exactly is...

May 16, 2026 · 4 min read

A Harness Engineering Best Practices case completed in collaboration with Codex. Today, the friends saw that

A Harness Engineering Best Practices case completed in collaboration with Codex. Today, the friends saw that Codex was about to reset the quota again, so they adjusted...

May 16, 2026 · 2 min read

A new way to play "Codex app" on mobile! The mobile "Codex app" essentially does not ha...

A new way to play "Codex app" on mobile! The mobile "Codex app" essentially does not have a separate development environment, it is just the Codex remote console in the ChatGPT...

May 16, 2026 · 2 min read

Is Grep All You Need? If you are still thinking of “rag = Vector Database” as the defau...

Is Grep All You Need? If you are still thinking of “rag = Vector Database” as the default answer, then I recommend that you read this paper. It is a bit counterintuitive: the au...

May 16, 2026 · 3 min read

The most interesting thing about Tencent's Agent Memory open source project is that it uses a lightweight symbol-like traversal method to manage complex memory relationships.

The most interesting thing about Tencent's Agent Memory open source project is that it uses a lightweight symbol-like traversal method to manage complex memory relationships. Th...

May 15, 2026 · 2 min read

CRM is degenerating into infrastructure. What is really valuable is the AI orchestratio...

CRM is degenerating into infrastructure. What is really valuable is the AI orchestration layer.⚡️ We have actually seen such a history once. At that time, Facebook had a friend...

May 15, 2026 · 1 min read

6-11: 6. Evals: LLM-as-judge + human evals Official Document | OpenAI Evaluation Best P...

6-11: 6. Evals: LLM-as-judge + human evals Official Document | OpenAI Evaluation Best Practices — Learn eval rubric, the applicable conditions for LLM-as-judge, the role of huma...

May 11, 2026 · 4 min read

Based on the recommendation of the boss, I combed through a list of high-quality AI Engineer learning materials, which is worth collecting and learning!

Based on the recommendation of the boss, I combed through a list of high-quality AI Engineer learning materials, which is worth collecting and learning! Too dry, too...

May 11, 2026 · 3 min read

AI Coding is now entering a very interesting phase. In the past, the most discussed top...

AI Coding is now entering a very interesting phase. In the past, the most discussed topics were model capabilities, context length, Agent Loop, Tool Use, and automated programmi...

May 10, 2026 · 2 min read

The Agent memory system is moving from "storing more and more" to "using more and more accurately".

The Agent memory system is moving from "storing more and more" to "using more and more accurately". Mem0's latest release, Memory Decay, is noteworthy. It addresses not...

May 9, 2026 · 2 min read

HTML might be the real interface of the agent era. Markdown is great for notes. But age...

HTML might be the real interface of the agent era. Markdown is great for notes. But agents do not just need notes. They need interfaces they can understand, manipulate, simulate...

May 9, 2026 · 2 min read

It is highly recommended that you read this long article written by Thariq, a core member of ClaudeCode: he brings an interesting perspective.

It is highly recommended that you read this long article written by Thariq, a core member of ClaudeCode: he brings an interesting perspective. HTML is the real working interface...

May 9, 2026 · 2 min read

During this time, I have been using the Codex App and the Codex CLI to develop iOS Apps at the same time

During this time, I have been using the Codex App and the Codex CLI to develop iOS Apps at the same time, especially in the SwiftUI + Xcode workflow, and it is becoming...

May 6, 2026 · 3 min read

Many people think that the difference between Codex CLI and Codex App is just: "One terminal, one GUI".

Many people think that the difference between Codex CLI and Codex App is just: "One terminal, one GUI". But once you really start using the computer, you will find that...

May 6, 2026 · 3 min read

Teacher Wu Yun’s introduction to codex goal is very detailed. I recommend everyone to r...

Teacher Wu Yun’s introduction to codex goal is very detailed. I recommend everyone to read it. At the same time, I will share a sample goal command that I verified and sent a fe...

May 4, 2026 · 2 min read

What is the limit of long Agent tasks? What I gained after running the codex goal for m...

What is the limit of long Agent tasks? What I gained after running the codex goal for more than ten hours. I have also thought about this problem before. Is a simple agent loop...

May 3, 2026 · 3 min read

Post something fun. What type of music is suitable for listening to when using codex im...

Post something fun. What type of music is suitable for listening to when using codex immersively? This is a playlist specially customized for vibe coders. This is the advice Cod...

May 2, 2026 · 4 min read

Using Codex goals well can truly unleash the potential of the GPT 5 model! OpenAI's rec...

Using Codex goals well can truly unleash the potential of the GPT 5 model! OpenAI's recently disclosed directions for Codex are obviously going in the following directions: - Lo...

May 1, 2026 · 6 min read

Many people underestimate a problem: the real danger of Agent is not the occasional wrong answer, but "systematic drift"

Many people discuss Agent, and their focus still remains on: - Is the answer correct this time - Is the tool adjusted accurately this time - Is the task completed this...

Apr 30, 2026 · 9 min read

Codex can actually not only do coding, but can also be used as a training partner! Try the following:

Codex can actually not only do coding, but can also be used as a training partner! Try the following: 1. First define "real mastery". Don't regard "understanding" as mastery. Fo...

Apr 29, 2026 · 4 min read

RLMs are the new inference models. Inference models are the first clear demonstration t...

RLMs are the new inference models. Inference models are the first clear demonstration that language model capabilities can be extended with test-time computation. Recursive lang...

Apr 21, 2026 · 2 min read

Just shipped: fireworks-tech-graph 🔥 Stop writing Mermaid DSL. Stop clicking around Ju...

Just shipped: fireworks-tech-graph 🔥 Stop writing Mermaid DSL. Stop clicking around Just describe your system in plain English → get a publication-ready SVG + PNG diagram in se...

Apr 10, 2026 · 1 min read

A short story about Cursor's counterattack. Cursor follows the path you mentioned - dee...

A short story about Cursor's counterattack. Cursor follows the path you mentioned - deeply distilling the programming scene to the extreme. The valuation in 2024 is several bill...

Apr 10, 2026 · 2 min read

Why does ClaudeCode use plain text + file system as its RAG system? Claude Code adopts a "plain text + file

Why does ClaudeCode use plain text + file system as its RAG system? Claude Code adopts a "plain text + file system" RAG design and completely abandons their...

Apr 10, 2026 · 2 min read

Fireworks-skill-memory major update! ! I received a lot of feedback after I posted it last time. I launched

Fireworks-skill-memory major update! ! I received a lot of feedback after I posted it last time. I launched v4 today and fixed several real problems I found after using...

Apr 5, 2026 · 3 min read

Why is this Harness open source project worth using? Claude Code's Skill ecosystem is r...

Why is this Harness open source project worth using? Claude Code's Skill ecosystem is rapidly expanding (official plugin marketplace + community skill). But "Claude can't rememb...

Mar 28, 2026 · 2 min read

A Harness engineering practice, I open sourced it! It’s quite interesting. After using...

A Harness engineering practice, I open sourced it! It’s quite interesting. After using Claude Code for two months, I discovered a crazy problem: it keeps making the same mistake...

Mar 27, 2026 · 2 min read

that's right Andrey, I'm an architect. Base on my expreince: The most difficult part of...

that's right Andrey, I'm an architect. Base on my expreince: The most difficult part of building software applications is orchestrating various services: account systems, paymen...

Mar 26, 2026 · 1 min read

Lin Junyang’s latest must-read masterpiece! AI has shifted from "reasoning thinking" to "agentic thinking".

Lin Junyang’s latest must-read masterpiece! AI has shifted from "reasoning thinking" to "agentic thinking". 🏀The frontier of competition is shifting from "better...

Mar 26, 2026 · 3 min read

Agency > Intelligence, Claw Agent is just the transition and beginning

A post that Andrej Karpathy reposted from Garry Tan (YC CEO) in February 2025 went viral some time ago. Elon Musk wasted no time and grabbed a sofa: This is Garry Tan’s original...

Mar 25, 2026 · 2 min read

Sora's departure is a good thing for OpenAI and AI entrepreneurs. Big companies do big things! I just woke up

Sora's departure is a good thing for OpenAI and AI entrepreneurs. Big companies do big things! I just woke up in the morning and saw that a group of friends had...

Mar 24, 2026 · 2 min read

Why do many claude code query tasks require the use of file search commands? Is there no other better way? I

Why do many claude code query tasks require the use of file search commands? Is there no other better way? I asked CC this question, and it gave me a very detailed answer...

Mar 23, 2026 · 2 min read

Harness Engineering: The next battlefield for AI engineers! AI Agent = Model + Harness....

Harness Engineering: The next battlefield for AI engineers! AI Agent = Model + Harness. If you're not a model, you're a Harness. OpenAI built 1 million lines of production code...

Mar 22, 2026 · 2 min read

🎬 Why I built Media Downloader for Claude Code — and why it's different from everything else out there.

🎬 Why I built Media Downloader for Claude Code — and why it's different from everything else out there. The problem: Getting stock images and videos is TEDIOUS. Open browser →...

Jan 20, 2026 · 2 min read

With this skill, you no longer need to manually search for picture websites, and no longer need to copy and paste links!

With this skill, you no longer need to manually search for picture websites, and no longer need to copy and paste links! ! Just say one sentence: "Help me download 10 xxx pictur...

Jan 20, 2026 · 3 min read

Experience sharing of using ob + claudian + palywright-skills + ob-skills + general-purpose-skills to generate visual logs

At one o'clock in the morning, I stared at the Canvas file just generated in Obsidian, feeling a little dazed. Figure 1: Automatically generated visual log...

Jan 14, 2026 · 10 min read

ClaudeCode Common Function Manual

📚 Core commands 📋 Command cheat sheet Command icon describes the frequency of use in one sentence /help🆘View the help document⭐⭐⭐⭐⭐ /clear🧹Clear the...

Jan 9, 2026 · 9 min read

Why is ClaudeCode so attractive?

One of my subjective feelings is not necessarily correct, let me give you some ideas: The biggest feeling is that coding can be well combined with AI through the...

Jan 9, 2026 · 2 min read

What will future travel and transportation look like based on AI? Such as Cybercab, Robovan and FSD

Musk’s blueprint for autonomous transportation, Cybercab, is a vehicle that completely eliminates the steering wheel and pedals and relies entirely on Tesla’s FSD...

Jan 7, 2026 · 12 min read

A historic moment for the world's first large model listed company. On January 8, 2026, as the world's first

A historic moment for the world's first large model listed company

Jan 7, 2026 · 17 min read

Manus Wide Research: 100+agent parallel research system Manus is the agent, achieving the most advanced performance in the GAIA benchmark test.

Manus Wide Research: 100+agent parallel research system Manus is the agent, achieving the most advanced performance in the GAIA benchmark test. Its Wide Research...

Jan 7, 2026 · 1 min read

With this Orange article, I will continue to share some content about ClaudeSkill and MCP to attract some traffic, haha.

With this Orange article, I will continue to share some content about ClaudeSkill and MCP to attract some traffic, haha. Analysis of the core features of Claude Skills....

Dec 29, 2025 · 5 min read

It seems that there will be many products and forms of agent in 2025. Why does it seem...

It seems that there will be many products and forms of agent in 2025. Why does it seem that only coding agent has emerged? One perspective is: Is AI Agent a technical term or a...

Dec 16, 2025 · 2 min read

Prompt: Role & Subject: A massive, encyclopedic 16:9 3D infographic poster titled "THE CHRONICLE OF [Person Name]".

Prompt: Role & Subject: A massive, encyclopedic 16:9 3D infographic poster titled "THE CHRONICLE OF [Person Name]". The visual style is a high-end fusion of...

Dec 7, 2025 · 2 min read

Why do I always recommend using Claude Code? Model capabilities are growing exponential...

Why do I always recommend using Claude Code? Model capabilities are growing exponentially (Exponential) 🚀 But the UX of development tools is still climbing linearly (Linear) 📷...

Nov 27, 2025 · 2 min read

Why do I always recommend using Claude Code? Model capabilities are growing exponential...

Why do I always recommend using Claude Code? Model capabilities are growing exponentially (Exponential) 🚀 But the UX of development tools is still climbing linearly (Linear) 🐢...

Nov 27, 2025 · 2 min read

Recently, the feeling of pushback brought by AI has become stronger and stronger, and sometimes it is even scary.

Recently, the feeling of pushback brought by AI has become stronger and stronger, and sometimes it is even scary. Think about it, it has been less than a year since...

Nov 27, 2025 · 4 min read

The Zara app is very interesting. You can control sliding to view web content just by recognizing gestures

The Zara app is very interesting. You can control sliding to view web content just by recognizing gestures with the camera. There are several points: 1. I recall that the female...

Nov 22, 2025 · 2 min read

This year I saw a lot of similar AI products focusing on memory or virtual companionship, which reminded me of a TV series from a few years ago.

This year I saw a lot of similar AI products focusing on memory or virtual companionship, which reminded me of a TV series from a few years ago. The general plot of "Upload to R...

Nov 20, 2025 · 2 min read

Claude recently added an official skill specifically designed to improve front-end interface design.

Claude recently added an official skill specifically designed to improve front-end interface design. I tried installing it and it worked really well. The screenshot is...

Nov 16, 2025 · 4 min read

Build the "Microservices Architecture Review Expert" Skill ❌ Inefficient way of writing: "You are an architect, help me review this microservice design."

Build the "Microservices Architecture Review Expert" Skill ❌ Inefficient way of writing: "You are an architect, help me review this microservice design." Problem: Too general...

Oct 20, 2025 · 2 min read

Breaking news about Gemini3 Gemini 3 has achieved significant performance improvements in many aspects compared to its predecessor Gemini 2.5 Pro.

Breaking news about Gemini3 Gemini 3 has achieved significant performance improvements in many aspects compared to its predecessor Gemini 2.5 Pro. Gemini 2.5 Pro is already...

Oct 14, 2025 · 2 min read

Large model training in 2023 is a simple two-step process: Pretraining is to throw massive data (books, web pages, codes, etc.)

Large model training in 2023 is a simple two-step process: Pretraining is to throw massive data (books, web pages, codes, etc.) to the model so that it can learn the...

Sep 27, 2025 · 1 min read

Large model training in 2023 is a simple two-step process: Pretraining is to throw massive data (books, web pages, codes, etc.)

Large model training in 2023 is a simple two-step process: Pretraining is to throw massive data (books, web pages, codes, etc.) to the model so that it can learn the...

Sep 27, 2025 · 1 min read

Claude finally takes action on the browser! I think this is indeed a good time. Claude...

Claude finally takes action on the browser! I think this is indeed a good time. Claude code has fully verified the coding capabilities based on its own model, coupled with a wel...

Aug 27, 2025 · 2 min read

Traditional large language models are often limited to short context windows such as 8K, 32K or 128K.

Traditional large language models are often limited to short context windows such as 8K, 32K or 128K. Expanding the context window to millions of tokens is not just a...

Aug 13, 2025 · 2 min read

RAG Technology Toolbox Layered Memory Retrieval: The simplest way to load all documents directly into the context window of LLM.

RAG Technology Toolbox Layered Memory Retrieval: The simplest way to load all documents directly into the context window of LLM. It is suitable for scenarios where the amount of...

Aug 10, 2025 · 2 min read

This article talks about multi-agent collaboration very well. I have recently observed...

This article talks about multi-agent collaboration very well. I have recently observed that starting from Calude’s multi-agent architecture article on building deep research, to...

Aug 7, 2025 · 2 min read

ClaudeCode best practices: Part 1: Provide clear context just like communicating with people.

ClaudeCode best practices: Part 1: Provide clear context just like communicating with people. First of all, the first and most important core concept shared by the official is:...

Aug 4, 2025 · 4 min read

Ecosystem, community and business context For open source projects, factors other than technology, such as

Ecosystem, community and business context For open source projects, factors other than technology, such as community health, business support and licensing agreements...

Jul 28, 2025 · 2 min read

Backend Technologies: Python/Flask vs. GoDify: Its back-end API is mainly built in Pyth...

Backend Technologies: Python/Flask vs. GoDify: Its back-end API is mainly built in Python language and uses lightweight Flask as the web framework. By analyzing its pyproject.to...

Jul 27, 2025 · 2 min read

Memory Is Not Context

Useful memory is not what survives by default. It is what a system learns to preserve on purpose.

April 2026 · 12 min read

Field Notes

Full ledger

Live judgment for this desk is filed by date in the field-notes ledger; enter through the full ledger by month.