bairui blog
为了根除“AI 味”,我翻回了一百年前的语言学
从欧化中文、现代汉语和 AI 写作痕迹谈起,借助语言学视角重新理解怎样写出更自然、更有人的中文表达。
8918 words
|
45 minutes
Cover Image of the Post
让 Context 流动起来:Agent 数字分身如何打通个人工作流与多人协作
从 Context Engineering 的视角复盘数字分身 Agent 的半年实践,讨论个人知识库、飞书工作流、多人协作和脱敏发布链路。
6550 words
|
33 minutes
Cover Image of the Post
万字长文:从 Spec Coding 到 Harness,AI Coding 的两次范式转变与实践
回顾 Vibe Coding、Spec Coding 到 Harness Engineering 的范式演进,讨论 AI Coding 工具链与约束工程的实践价值。
8750 words
|
44 minutes
Cover Image of the Post
Agent的实现逻辑(OpenClaw)
基于 OpenClaw 的抓包、日志与运行配置,拆解 Agent 消息处理链路、并发队列、会话存储、工具注入和上下文消耗。
3981 words
|
20 minutes
Cover Image of the Post
AI Agent评测入门
面向入门读者介绍 Agent 评测的基本概念、离线与在线评估、人工评测、评测集建设和核心指标设计。
4245 words
|
21 minutes
Cover Image of the Post
别再靠人肉评价 Agent 了:EvalAgent——客观、精准、高效评测的落地指南
围绕 EvalAgent 的端到端评测实践,拆解自动化评分、Trace 取证、视觉产物检测和多维质量判断的落地方法。
9004 words
|
45 minutes
Cover Image of the Post
Agent评测方法论梳理
系统梳理 Agent 评测的结果、过程、安全与成本框架,并讨论 Benchmark、Trace 与 Agent 治理的落地路径。
17809 words
|
89 minutes
Cover Image of the Post
大模型评测经验总结与探讨
从 QA 视角梳理大模型评测的价值定位、阶段目标、评测体系、安全底线与业务决策支撑方法。
6194 words
|
31 minutes
Cover Image of the Post