Agents & InferenceSimon Willison

llm 0.33

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

The llm library and CLI have upgraded to OpenAI Python SDK 3.x and httpx2, introducing stateless, per-call API key injection for embedding models and collections. This allows concurrent, multi-tenant embedding pipelines to run safely without mutating shared model state. Additionally, new template chaining capabilities enable developers to separate and combine runtime model configurations with prompt templates on the fly.

LinkedIn

Two AI summaries of each story, blind-voted — see today's agents & inference digest →