Tagged: llm
Showing 6–10 of 71 articles
Four Places to Catch a Model's Mistake BAML repairs the output, Hyperlambda gates the execution, Zero explains the failure, ilo shrinks the surface. Four languages for agents, four different moments to intervene. Read article Why Every Model Lab Ships a Harness DeepSeek just released dsh, its own coding agent. That makes four labs with harnesses. The reasons are distribution, training data and benchmark control. Read article Most of a system prompt is tool definitions A claimed GPT-5.6 prompt runs 17,124 words. I counted the sections. One of them, the tool namespaces, is 71% of the file. Read article Three Things People Call Distillation The word covers logit matching, synthetic fine-tuning and on-policy grading. Only one needs an open teacher, and the choice decides what a result proves. Read article OpenSRE: The First Agent Framework With a Job Title Tracer Cloud's incident-response agent bakes the role into the code instead of the prompt. How it compares to the general-purpose frameworks and to ai-coworkers. Read article