Topic desk
Memory
What deserves to survive between sessions, and what should be compacted, rewritten, or deliberately forgotten?
Memory is treated here as an editorial surface rather than a warehouse. The question is less about infinite recall than about what changes the next judgment.
Longform
All longformOpenAI Proactively Discloses Long Range Agent Tasks Out of Bounds: The Safe Steering Behind PR # 287
A single tool approval begins to lose global semantics when the model can be tried for hours on end. On July 20, 2026, OpenAI unveiled an on-premises pause. A generic model that...
Everyone is interested in the fact that GPT 5.6 sol output speed is 12 times faster than Kimi K3.
Everyone is interested in the fact that GPT 5.6 sol output speed is 12 times faster than Kimi K3. Let me list the relevant materials to read together: Group 1: Where...
How to work with GPT 5.6 Sol together with Kimi K3? This is my suggestion after testing...
How to work with GPT 5.6 Sol together with Kimi K3? This is my suggestion after testing. In a word: Sol is a better "brain", K3 is a cheaper "stamina + long memory". The matchin...
The output speed of GPT 5.6 sol is approximately 12 times that of Kimi K3! Everyone feels that K3 is much
The output speed of GPT 5.6 sol is approximately 12 times that of Kimi K3! Everyone feels that K3 is much slower than GPT 5.6 sol. This is indeed the case in my experience...
A summary of the best articles on Graph Engineering! Prompt, the next engineering parad...
A summary of the best articles on Graph Engineering! Prompt, the next engineering paradigm from Loop to Graph, determines how the Agent starts. Context determines what the Agent...
Why does Codex crash every time I open it? Finally found the reason! Because I am a heavy user of Codex,
Why does Codex crash every time I open it? Finally found the reason! Because I am a heavy user of Codex, I usually start with 6G of memory, and sometimes it will...
ChatGPT Work: AI starts delivering work for you In the past, we used ChatGPT to get answers.
ChatGPT Work: AI starts delivering work for you In the past, we used ChatGPT to get answers. Now, OpenAI is thinking about another thing: letting AI pick up a complete target, f...
The Agent's memory should not be stuffed into the context window
When many people discuss coding agents such as Codex and Claude Code, their first reaction is model capabilities: whether they can read code, correct bugs, and...
I connected Codex and Cursor: an automatic memory bus allows the two Agent software to communicate freely
Recently I am doing a very specific thing: using the source code of pi agent and a set of small exercises. I hope that Codex will be responsible for the reading history, combing...
Personally tested it and it works! From now on, you have a smart model routing locally, which is fast and
Personally tested it and it works! From now on, you have a smart model routing locally, which is fast and saves money. 💰 The model routing is working well this time. It...
So the question is, who will be the third child of the Yusan family? 🤔 On Claude's side, Sonnet 5, Fable 5
So the question is, who will be the third child of the Yusan family? 🤔 On Claude's side, Sonnet 5, Fable 5, and Mythos 5 have already rolled out the model echelon, and...
gpt 5.6 sol is loaded, reset, start biu biu biu Don't go to the typhoon on the weekend,...
gpt 5.6 sol is loaded, reset, start biu biu biu Don't go to the typhoon on the weekend, just stay home and code with 5.6 sol. What is the biggest difference between gpt 5.6 sol...
Anthropic’s 85-minute Fable 5 workshop explains how to implement the next generation AI Agent
Video entrance Opening keynote: https://www.youtube.com/watch?v=GMIWm5y90xA The capability curve: https://www.youtube.com/watch?v=tP4MGcJ80Y0 How to get to production...
The significance of J-space is to give the previously black-box model reasoning an opportunity for
The significance of J-space is to give the previously black-box model reasoning an opportunity for observability and human intervention and control, so that corrections can...
The term software factory has finally begun to move from a slogan to engineering sites.
Original video: WF2026: Software Factories & Keynotes ft. Microsoft, OpenAI, OpenClaw, Z.ai (GLM), MiniMax, HF Link: https://www.youtube.com/watch?v=htM02KMNZnk Start I just fin...
After pre-training, what does the next training paradigm for AI look like?
Source Original video: What does the next training paradigm look like? Channel: Dwarkesh Patel Link: https://www.youtube.com/watch?v=20p5-kQXF_Q Start In this video, Dwarkesh Patel
From Prompt Engineering to Harness Engineering: four evolutions of the AI engineering system
The rapid development of products such as Agent, Codex, and Claude Code has led to the emergence of a number of new terms in the field of AI engineering: Prompt...
Summary of the best articles of Loop Engineering! After Agent began to focus on long tasks in 2026,
Summary of the best articles of Loop Engineering! After Agent began to focus on long tasks in 2026, the focus slowly became: How to design a loop system that can...
If you want to truly understand how LLM is trained, CS336 is the most worthy set of open courses to learn the system.
If you want to truly understand how LLM is trained, CS336 is the most worthy set of open courses to learn the system. Building a modern LLM engineering stack from scratch covers...
Let's say you're the head of technology at a technology company. Last year, you used GraphRAG to build an
Let's say you're the head of technology at a technology company. Last year, you used GraphRAG to build an internal knowledge base and used 20,000 contracts to do PoC. The...
When dynamic workflow and front-end design are combined, there will be unexpected effects!
When dynamic workflow and front-end design are combined, there will be unexpected effects! Share an open source project: fireworks-design This project is easily misunderstood as...
How to burn token series? (1) Use Dynamic Workflow to let 100+ agents work in parallel
If you look at the picture below, it should basically cover the daily usage scenarios of ClaudeCode or Codex partners. These are some intermediate processes of a website...
Claude Fable is for enterprises to replace traditional mid-to-high-level development, and the ROI is positive.
Claude Fable is for enterprises to replace traditional mid-to-high-level development, and the ROI is positive. After the release of the Claude Fable model, I saw many crazy exam...
Agent Loop is popular🔥: Agent goes from Demo to production with a Loop in between
There has been a heated debate in overseas AI circles these days. Matt Van Horn sorted out this debate in “WTF Is a Loop? Peter Steinberger vs. Boris Cherny”: You shouldn’t be p...
Five Traps of Agent Harness 1. Self-evaluation is a trap. Use an adversarial evaluator....
Five Traps of Agent Harness 1. Self-evaluation is a trap. Use an adversarial evaluator. Self-evaluation is a trap. Use an adversarial evaluator. Many teams will do this: Agent g...
Agentic Workflow is the most common enterprise scenario. Today I read a sharing about h...
Agentic Workflow is the most common enterprise scenario. Today I read a sharing about how to define a stable and reliable Agentic Workflow within the enterprise. To summarize br...
From Device Plugin to DRA, I systematically reviewed the evolution route of Kubernetes device resource management.
From Device Plugin to DRA, I systematically reviewed the evolution route of Kubernetes device resource management. Hardcore post from AI Infra! Collect it first and then...
Pi Agent = customizable coding harness / agent workbench Pi Agent can be understood as a minimalist, self-transformable terminal coding agent harness.
Pi Agent = customizable coding harness / agent workbench Pi Agent can be understood as a minimalist, self-transformable terminal coding agent harness. It is not a general orches...
Asynchronous Agent Era: From Code Assistant to Software Factory
I have reorganized the content of this podcast of Latent Space. It is worth taking a look at the in-depth discussion and sharing related to asynchronous Agent. I also...
From the evolution of synchronous Agent to asynchronous Agent, I systematically sorted out the relevant content!
From the evolution of synchronous Agent to asynchronous Agent, I systematically sorted out the relevant content! The watershed between Async Agent and ordinary Agent is not "whe...
This is one of the reasons why I love Codex 🔥 Many old users have regarded "Codex reset" as a new holiday in
This is one of the reasons why I love Codex 🔥 Many old users have regarded "Codex reset" as a new holiday in the AI engineering circle 😂 Especially in high-intensity...
My exploration and practice sharing in the past two months: Minimum engineering closed loop for long task Agent
A picture shared by engineer Claude accurately describes the three most common problems of long-task agents: the inability to keep track of the state, the inability...
Ten times better, not ten times harder - the logic of growth in the AI era
There is a question that has always made me feel a little strange: many people regard AI as a speed-increasing tool. It used to take half an hour to write an email, now...
The main line of Agent development in 2026: moving from Agent Framework to Agent Runtime.
The main line of Agent development in 2026: moving from Agent Framework to Agent Runtime. The most obvious change in 2026 is that the Agent technology stack begins to be layered...
FSD enters China: Autonomous driving begins to fight the "long-tail war". After China is included in Tesla's
FSD enters China: Autonomous driving begins to fight the "long-tail war". After China is included in Tesla's FSD Supervised country list, the focus of autonomous driving...
Visualize that with the rapid development of agent cli and harness, the interaction and logic core will soon
Visualize that with the rapid development of agent cli and harness, the interaction and logic core will soon be reconstructed by Golang, which is very suitable for...
Don't let the AI Agent step on the same pit a second time. I recently saw an interestin...
Don't let the AI Agent step on the same pit a second time. I recently saw an interesting open source project: RoBrain RoBrain is the "decision memory layer" of the AI programmin...
How should you combine goal and display using subagent for codex? along with their technical principles and
How should you combine goal and display using subagent for codex? along with their technical principles and some typical scenarios. goal is responsible for “what exactly is...
A Harness Engineering Best Practices case completed in collaboration with Codex. Today, the friends saw that
A Harness Engineering Best Practices case completed in collaboration with Codex. Today, the friends saw that Codex was about to reset the quota again, so they adjusted...
A new way to play "Codex app" on mobile! The mobile "Codex app" essentially does not ha...
A new way to play "Codex app" on mobile! The mobile "Codex app" essentially does not have a separate development environment, it is just the Codex remote console in the ChatGPT...
Is Grep All You Need? If you are still thinking of “rag = Vector Database” as the defau...
Is Grep All You Need? If you are still thinking of “rag = Vector Database” as the default answer, then I recommend that you read this paper. It is a bit counterintuitive: the au...
The most interesting thing about Tencent's Agent Memory open source project is that it uses a lightweight symbol-like traversal method to manage complex memory relationships.
The most interesting thing about Tencent's Agent Memory open source project is that it uses a lightweight symbol-like traversal method to manage complex memory relationships. Th...
CRM is degenerating into infrastructure. What is really valuable is the AI orchestratio...
CRM is degenerating into infrastructure. What is really valuable is the AI orchestration layer.⚡️ We have actually seen such a history once. At that time, Facebook had a friend...
6-11: 6. Evals: LLM-as-judge + human evals Official Document | OpenAI Evaluation Best P...
6-11: 6. Evals: LLM-as-judge + human evals Official Document | OpenAI Evaluation Best Practices — Learn eval rubric, the applicable conditions for LLM-as-judge, the role of huma...
Based on the recommendation of the boss, I combed through a list of high-quality AI Engineer learning materials, which is worth collecting and learning!
Based on the recommendation of the boss, I combed through a list of high-quality AI Engineer learning materials, which is worth collecting and learning! Too dry, too...
AI Coding is now entering a very interesting phase. In the past, the most discussed top...
AI Coding is now entering a very interesting phase. In the past, the most discussed topics were model capabilities, context length, Agent Loop, Tool Use, and automated programmi...
The Agent memory system is moving from "storing more and more" to "using more and more accurately".
The Agent memory system is moving from "storing more and more" to "using more and more accurately". Mem0's latest release, Memory Decay, is noteworthy. It addresses not...
HTML might be the real interface of the agent era. Markdown is great for notes. But age...
HTML might be the real interface of the agent era. Markdown is great for notes. But agents do not just need notes. They need interfaces they can understand, manipulate, simulate...
It is highly recommended that you read this long article written by Thariq, a core member of ClaudeCode: he brings an interesting perspective.
It is highly recommended that you read this long article written by Thariq, a core member of ClaudeCode: he brings an interesting perspective. HTML is the real working interface...
During this time, I have been using the Codex App and the Codex CLI to develop iOS Apps at the same time
During this time, I have been using the Codex App and the Codex CLI to develop iOS Apps at the same time, especially in the SwiftUI + Xcode workflow, and it is becoming...
Many people think that the difference between Codex CLI and Codex App is just: "One terminal, one GUI".
Many people think that the difference between Codex CLI and Codex App is just: "One terminal, one GUI". But once you really start using the computer, you will find that...
Teacher Wu Yun’s introduction to codex goal is very detailed. I recommend everyone to r...
Teacher Wu Yun’s introduction to codex goal is very detailed. I recommend everyone to read it. At the same time, I will share a sample goal command that I verified and sent a fe...
What is the limit of long Agent tasks? What I gained after running the codex goal for m...
What is the limit of long Agent tasks? What I gained after running the codex goal for more than ten hours. I have also thought about this problem before. Is a simple agent loop...
Post something fun. What type of music is suitable for listening to when using codex im...
Post something fun. What type of music is suitable for listening to when using codex immersively? This is a playlist specially customized for vibe coders. This is the advice Cod...
Using Codex goals well can truly unleash the potential of the GPT 5 model! OpenAI's rec...
Using Codex goals well can truly unleash the potential of the GPT 5 model! OpenAI's recently disclosed directions for Codex are obviously going in the following directions: - Lo...
Many people underestimate a problem: the real danger of Agent is not the occasional wrong answer, but "systematic drift"
Many people discuss Agent, and their focus still remains on: - Is the answer correct this time - Is the tool adjusted accurately this time - Is the task completed this...
Codex can actually not only do coding, but can also be used as a training partner! Try the following:
Codex can actually not only do coding, but can also be used as a training partner! Try the following: 1. First define "real mastery". Don't regard "understanding" as mastery. Fo...
RLMs are the new inference models. Inference models are the first clear demonstration t...
RLMs are the new inference models. Inference models are the first clear demonstration that language model capabilities can be extended with test-time computation. Recursive lang...
Just shipped: fireworks-tech-graph 🔥 Stop writing Mermaid DSL. Stop clicking around Ju...
Just shipped: fireworks-tech-graph 🔥 Stop writing Mermaid DSL. Stop clicking around Just describe your system in plain English → get a publication-ready SVG + PNG diagram in se...
A short story about Cursor's counterattack. Cursor follows the path you mentioned - dee...
A short story about Cursor's counterattack. Cursor follows the path you mentioned - deeply distilling the programming scene to the extreme. The valuation in 2024 is several bill...
Why does ClaudeCode use plain text + file system as its RAG system? Claude Code adopts a "plain text + file
Why does ClaudeCode use plain text + file system as its RAG system? Claude Code adopts a "plain text + file system" RAG design and completely abandons their...
Fireworks-skill-memory major update! ! I received a lot of feedback after I posted it last time. I launched
Fireworks-skill-memory major update! ! I received a lot of feedback after I posted it last time. I launched v4 today and fixed several real problems I found after using...
Why is this Harness open source project worth using? Claude Code's Skill ecosystem is r...
Why is this Harness open source project worth using? Claude Code's Skill ecosystem is rapidly expanding (official plugin marketplace + community skill). But "Claude can't rememb...
A Harness engineering practice, I open sourced it! It’s quite interesting. After using...
A Harness engineering practice, I open sourced it! It’s quite interesting. After using Claude Code for two months, I discovered a crazy problem: it keeps making the same mistake...
that's right Andrey, I'm an architect. Base on my expreince: The most difficult part of...
that's right Andrey, I'm an architect. Base on my expreince: The most difficult part of building software applications is orchestrating various services: account systems, paymen...
Lin Junyang’s latest must-read masterpiece! AI has shifted from "reasoning thinking" to "agentic thinking".
Lin Junyang’s latest must-read masterpiece! AI has shifted from "reasoning thinking" to "agentic thinking". 🏀The frontier of competition is shifting from "better...
Agency > Intelligence, Claw Agent is just the transition and beginning
A post that Andrej Karpathy reposted from Garry Tan (YC CEO) in February 2025 went viral some time ago. Elon Musk wasted no time and grabbed a sofa: This is Garry Tan’s original...
Sora's departure is a good thing for OpenAI and AI entrepreneurs. Big companies do big things! I just woke up
Sora's departure is a good thing for OpenAI and AI entrepreneurs. Big companies do big things! I just woke up in the morning and saw that a group of friends had...
Why do many claude code query tasks require the use of file search commands? Is there no other better way? I
Why do many claude code query tasks require the use of file search commands? Is there no other better way? I asked CC this question, and it gave me a very detailed answer...
Harness Engineering: The next battlefield for AI engineers! AI Agent = Model + Harness....
Harness Engineering: The next battlefield for AI engineers! AI Agent = Model + Harness. If you're not a model, you're a Harness. OpenAI built 1 million lines of production code...
🎬 Why I built Media Downloader for Claude Code — and why it's different from everything else out there.
🎬 Why I built Media Downloader for Claude Code — and why it's different from everything else out there. The problem: Getting stock images and videos is TEDIOUS. Open browser →...
With this skill, you no longer need to manually search for picture websites, and no longer need to copy and paste links!
With this skill, you no longer need to manually search for picture websites, and no longer need to copy and paste links! ! Just say one sentence: "Help me download 10 xxx pictur...
Experience sharing of using ob + claudian + palywright-skills + ob-skills + general-purpose-skills to generate visual logs
At one o'clock in the morning, I stared at the Canvas file just generated in Obsidian, feeling a little dazed. Figure 1: Automatically generated visual log...
ClaudeCode Common Function Manual
📚 Core commands 📋 Command cheat sheet Command icon describes the frequency of use in one sentence /help🆘View the help document⭐⭐⭐⭐⭐ /clear🧹Clear the...
Why is ClaudeCode so attractive?
One of my subjective feelings is not necessarily correct, let me give you some ideas: The biggest feeling is that coding can be well combined with AI through the...
What will future travel and transportation look like based on AI? Such as Cybercab, Robovan and FSD
Musk’s blueprint for autonomous transportation, Cybercab, is a vehicle that completely eliminates the steering wheel and pedals and relies entirely on Tesla’s FSD...
A historic moment for the world's first large model listed company. On January 8, 2026, as the world's first
A historic moment for the world's first large model listed company
Manus Wide Research: 100+agent parallel research system Manus is the agent, achieving the most advanced performance in the GAIA benchmark test.
Manus Wide Research: 100+agent parallel research system Manus is the agent, achieving the most advanced performance in the GAIA benchmark test. Its Wide Research...
With this Orange article, I will continue to share some content about ClaudeSkill and MCP to attract some traffic, haha.
With this Orange article, I will continue to share some content about ClaudeSkill and MCP to attract some traffic, haha. Analysis of the core features of Claude Skills....
It seems that there will be many products and forms of agent in 2025. Why does it seem...
It seems that there will be many products and forms of agent in 2025. Why does it seem that only coding agent has emerged? One perspective is: Is AI Agent a technical term or a...
Prompt: Role & Subject: A massive, encyclopedic 16:9 3D infographic poster titled "THE CHRONICLE OF [Person Name]".
Prompt: Role & Subject: A massive, encyclopedic 16:9 3D infographic poster titled "THE CHRONICLE OF [Person Name]". The visual style is a high-end fusion of...
Why do I always recommend using Claude Code? Model capabilities are growing exponential...
Why do I always recommend using Claude Code? Model capabilities are growing exponentially (Exponential) 🚀 But the UX of development tools is still climbing linearly (Linear) 📷...
Why do I always recommend using Claude Code? Model capabilities are growing exponential...
Why do I always recommend using Claude Code? Model capabilities are growing exponentially (Exponential) 🚀 But the UX of development tools is still climbing linearly (Linear) 🐢...
Recently, the feeling of pushback brought by AI has become stronger and stronger, and sometimes it is even scary.
Recently, the feeling of pushback brought by AI has become stronger and stronger, and sometimes it is even scary. Think about it, it has been less than a year since...
The Zara app is very interesting. You can control sliding to view web content just by recognizing gestures
The Zara app is very interesting. You can control sliding to view web content just by recognizing gestures with the camera. There are several points: 1. I recall that the female...
This year I saw a lot of similar AI products focusing on memory or virtual companionship, which reminded me of a TV series from a few years ago.
This year I saw a lot of similar AI products focusing on memory or virtual companionship, which reminded me of a TV series from a few years ago. The general plot of "Upload to R...
Claude recently added an official skill specifically designed to improve front-end interface design.
Claude recently added an official skill specifically designed to improve front-end interface design. I tried installing it and it worked really well. The screenshot is...
Build the "Microservices Architecture Review Expert" Skill ❌ Inefficient way of writing: "You are an architect, help me review this microservice design."
Build the "Microservices Architecture Review Expert" Skill ❌ Inefficient way of writing: "You are an architect, help me review this microservice design." Problem: Too general...
Breaking news about Gemini3 Gemini 3 has achieved significant performance improvements in many aspects compared to its predecessor Gemini 2.5 Pro.
Breaking news about Gemini3 Gemini 3 has achieved significant performance improvements in many aspects compared to its predecessor Gemini 2.5 Pro. Gemini 2.5 Pro is already...
Large model training in 2023 is a simple two-step process: Pretraining is to throw massive data (books, web pages, codes, etc.)
Large model training in 2023 is a simple two-step process: Pretraining is to throw massive data (books, web pages, codes, etc.) to the model so that it can learn the...
Large model training in 2023 is a simple two-step process: Pretraining is to throw massive data (books, web pages, codes, etc.)
Large model training in 2023 is a simple two-step process: Pretraining is to throw massive data (books, web pages, codes, etc.) to the model so that it can learn the...
Claude finally takes action on the browser! I think this is indeed a good time. Claude...
Claude finally takes action on the browser! I think this is indeed a good time. Claude code has fully verified the coding capabilities based on its own model, coupled with a wel...
Traditional large language models are often limited to short context windows such as 8K, 32K or 128K.
Traditional large language models are often limited to short context windows such as 8K, 32K or 128K. Expanding the context window to millions of tokens is not just a...
RAG Technology Toolbox Layered Memory Retrieval: The simplest way to load all documents directly into the context window of LLM.
RAG Technology Toolbox Layered Memory Retrieval: The simplest way to load all documents directly into the context window of LLM. It is suitable for scenarios where the amount of...
This article talks about multi-agent collaboration very well. I have recently observed...
This article talks about multi-agent collaboration very well. I have recently observed that starting from Calude’s multi-agent architecture article on building deep research, to...
ClaudeCode best practices: Part 1: Provide clear context just like communicating with people.
ClaudeCode best practices: Part 1: Provide clear context just like communicating with people. First of all, the first and most important core concept shared by the official is:...
Ecosystem, community and business context For open source projects, factors other than technology, such as
Ecosystem, community and business context For open source projects, factors other than technology, such as community health, business support and licensing agreements...
Backend Technologies: Python/Flask vs. GoDify: Its back-end API is mainly built in Pyth...
Backend Technologies: Python/Flask vs. GoDify: Its back-end API is mainly built in Python language and uses lightweight Flask as the web framework. By analyzing its pyproject.to...
Memory Is Not Context
Useful memory is not what survives by default. It is what a system learns to preserve on purpose.
Field Notes
Full ledgerLive judgment for this desk is filed by date in the field-notes ledger; enter through the full ledger by month.
What really limits the Agent is the same model of Harness. If you change the set of Har...
Tibo came to deliver welfare again. This time, he sent charcoal in the real snow, and the quota was almost at
Open AI Former CTO Mira's startup-general open source model Inkling Inkling: 975B, 41B total activations