AI Kai
文章关于合作
English中文
首页/来自构建循环的笔记。

文章

来自构建循环的笔记。

面向使用 Codex、Claude Code、Cursor 和各类 Agent 的开发者,分享实测过的 AI 编程工作流、编程 Agent 现场笔记和真实的大模型评测。

DeepSeek Harness walkthrough
Agent WorkflowsAug 23, 2026

Inside DeepSeek Harness and Its Plugin-Based Agent Runtime

DeepSeek Harness separates models, tools, sessions, permissions, and interfaces into plugins that builders can replace and combine.

阅读文章
DeepSeek V4 Pro benchmark results
AI Dev ToolsAug 23, 2026

DeepSeek V4 Flash Review: Where It Fits in an Agent Workflow

A hands-on look at DeepSeek V4 Flash 0731, its vendor-reported agent benchmarks, and the coding tasks where I would use it.

阅读文章
GPT-5.6 model picker
CodexAug 23, 2026

How I Choose GPT-5.6 Sol, Terra, Luna, and Effort Levels

A practical decision guide for choosing a GPT-5.6 model and reasoning effort level based on task difficulty, speed, and iteration cost.

阅读文章
LoopX local control plane interface
Agent WorkflowsAug 23, 2026

LoopX: A Local Control Plane for Long-Running Agent Work

LoopX keeps goals, evidence, ownership, quotas, and handoffs stable across long-running agent sessions without replacing the agent runtime.

阅读文章
Claude Code session optimization tips
Claude CodeAug 23, 2026

How to Get More Value from Claude Code Sessions

Practical Claude Code habits that reduce context noise, protect prompt cache reuse, and keep recurring work in smaller sessions.

阅读文章
ModLens vision demo for a text-only coding model
AI Dev ToolsAug 23, 2026

Give a Text-Only Coding Agent Vision with ModLens

Use ModLens to route screenshots and image files through a configured vision engine, then return structured evidence to a text-only coding agent.

阅读文章
Qwen3.8-27B running locally on a Mac
AI Dev ToolsAug 23, 2026

Run Qwen3.8-27B Locally on Apple Silicon

Install an Unsloth GGUF build of Qwen3.8-27B with Ollama or llama.cpp, then choose a quant that leaves enough unified memory for the runtime.

阅读文章
AI Dev ToolsJul 29, 2026

LLM Inference Handbook: the reference I wish existed a year ago

Same model, different inference stack, wildly different latency and cost. Modular's handbook covers VRAM math, KV cache, batching, and serving.

阅读文章
CodexJul 26, 2026

Codex Just Shipped Two Features That Matter

A practical look at ChatGPT Voice and multi-folder projects in Codex, including their workflow benefits, limits, and setup details.

阅读文章
Agent WorkflowsJul 26, 2026

Can Kimi K3 Handle a Real Production Backend?

A hands-on Kimi K3 backend audit across 68K lines of FastAPI code, including query fixes, measured results, and the full task cost.

阅读文章
上一页
123456
下一页

Kai 为用 AI 交付的人而写。

隐私媒体资料
XGitHubLinkedInEmail