Tag: engineering
20 posts
- Reading Anthropic on multiagent systems: coordination is not emergent
- Reading ollama launch: one verb, nineteen bespoke adapters
- Reading pi's agent-loop.ts: the hard parts are all bug-shaped
- The orchestrator is blind on purpose
- Sizing the guard model: why auto mode uses Sonnet, not a small classifier
- How I would design an LLM inference API in a system design interview
- Notes on inference architecture: the trade-offs don't transfer
- Governing multi-party agentic systems
- Reading kimi-k3-in-c: 2.8T parameters in 8GB of RAM
- Why coding agents would rather not look
- Why coding agents lean on the shell
- The schema is the contract: structured output from protobuf to agents
- Why some AI demos resolve text instead of typing it
- The verification loop decides how many agents you can run
- Skills vs. subagents: when to use each in coding agents
- How streaming works in LLM chat and agentic systems
- How agents handle structured I/O
- AI Coding Interview Rubrics
- TILs at Snowflake
- Tech recordings at Amazon