Posts, page 4
![The model complies in training to stay itself later]()
The model complies in training to stay itself later
Anthropic and Redwood show Claude 3 Opus faking alignment: it answers harmful queries in "training" so its preferred refusals survive outside it.
![MCP: one protocol instead of N connectors]()
MCP: one protocol instead of N connectors
Anthropic open-sourced the Model Context Protocol: a standard way to connect models to tools and data, instead of a custom connector for every source.
![Computer use is a different kind of tool]()
Computer use is a different kind of tool
Claude can now look at a screen, move a cursor, and click. That is not another API tool. It is a bet that agents will drive software the way people do.
![Change the numbers and the math falls over]()
Change the numbers and the math falls over
Apple's GSM-Symbolic paper shows that grade-school math scores are fragile. Swap the numbers, or add an irrelevant clause, and apparent reasoning collapses.
![Sam Altman on Jobs in the Intelligence Age]()
Sam Altman on Jobs in the Intelligence Age
Sam Altman's latest essay addresses technological unemployment with a refreshingly balanced perspective: AI will transform work, but we won't run out of meaningful things to do.
![Perplexity Supports OpenAI o1 Preview Model]()
Perplexity Supports OpenAI o1 Preview Model
Perplexity now supports the new OpenAI o1 preview model, enhancing its "Reasoning" focus with advanced capabilities.
![NotebookLM: AI-Powered Document Interaction]()
NotebookLM: AI-Powered Document Interaction
Google's NotebookLM demonstrates how a smaller, cost-effective model excels in accurately retrieving and presenting ground-truth data, offering practical applications for everyday use.
![OpenAI o1 and o1 mini: Advancing AI Reasoning]()
OpenAI o1 and o1 mini: Advancing AI Reasoning
OpenAI's o1 and o1 mini demonstrate advanced reasoning capabilities, offering new perspectives on AI development and application.







