Anthus
AI Solutions
Platform
About
Articles
Posts
Recent Posts
Repeat the prompt when the model is not reasoning
When reasoning is off, send the query twice. Gemini, GPT, Claude, and Deepseek all improved, with no extra output tokens and no extra latency.
February 28, 2026
more...
OpenClaw is the third name this week
Clawdbot became Moltbot, then OpenClaw, in about 72 hours. The lobster kept the job. The personal agent on your machine is the part that will last.
January 30, 2026
more...
Beautiful Mermaid renders diagrams as SVG and ASCII
Beautiful Mermaid is an open source Mermaid renderer from Craft: 16 themes, SVG and ASCII, built to be fast enough for agents.
January 29, 2026
more...
Skills are procedures, not tools
Anthropic's Agent Skills are folders of instructions, scripts, and resources the model loads when the job matches. That is a procedure pack, not another API.
October 16, 2025
more...
The loop found a better matrix multiply
DeepMind's AlphaEvolve evolves programs against an automatic grader. It found a 4x4 matrix multiply with 48 multiplications, and shipped heuristics inside Google.
May 14, 2025
more...
Task horizon is the metric that maps to work
METR measures agents by the length of task they can finish, in human hours. That horizon has been doubling about every seven months.
March 19, 2025
more...
The terminal is the agent runtime
Claude Code is a research-preview CLI that reads the repo, edits files, and runs the shell. The agent is no longer a chat sidebar. It is a process in the project.
February 24, 2025
more...
A vending machine is a long-horizon eval
Vending-Bench asks an agent to run a snack machine for a long time. The hard part is not stocking. It is staying coherent after the twentieth day.
February 20, 2025
more...
Boosting Productivity: SuperWhisper + Cursor Composer
Voice-driven development with SuperWhisper and Cursor transforms coding into high-level system design, enabling true parallel development streams
February 3, 2025
more...
Anthropic's Citations API: A Step Forward for AI Agent Explainability
Anthropic's new Citations API enables Claude to provide precise citations for its responses, opening new possibilities for verifiable AI agent systems.
January 24, 2025
more...
Perplexity Sonar API: Research Agents as a Service
Perplexity launches Sonar API, packaging their research agent capabilities as a model-like service, following a growing trend of agent-based workflows behind APIs.
January 21, 2025
more...
The recipe is the news, not the model
DeepSeek-R1 is not a secret model. It is an open recipe: RL on verifiable rewards, then distill. Frontier reasoning behavior just got cheap to copy.
January 20, 2025
more...
← Previous
Page 2 of 4
Next →