AI Agent Benchmark
Practical evaluation for agents in real workflows.
已形成评估方法草稿,继续收集真实工作流案例。
探索中
Practical evaluation for agents in real workflows.
已形成评估方法草稿,继续收集真实工作流案例。
Turn volunteer voice notes into useful reports.
从独立网页转向飞书工作流,工序已被复现。
See where attention goes across personal projects.
暂停迭代,时间线还没有成为日常使用习惯。
Sharing practical lessons on building with AI.
完成一次 45 分钟分享,尝试面向初学者的陪跑。