Seth Lazar

@sethlazar.org

Philosopher working on AI alignment, governance and adaptation Lab: https://mintresearch.org Self: https://sethlazar.org Newsletter: https://philosophyofcomputing.substack.com

Here's my contribution on using agents to support academic research. I've got a pipeline going now with coding agents that checks arxiv, twitter, bluesky, philpapers, a bunch of journals, many RSS feeds and more, classifies it against a long statement of my lab's interests...

🚨 New Study 🚨 @arxiv.bsky.social has recently decided to prohibit any 'position' paper from being submitted to its CS servers. Why? Because of the "AI slop", and allegedly higher ratios of LLM-generated content in review papers, compared to non-review papers.

And this is from Anthropic... We need to get LLMs out of learning contexts "Our findings suggest that AI-enhanced productivity is not a shortcut to competence and AI assistance should be carefully adopted into workflows to preserve skill formation..." arxiv.org/abs/2601.20245

How AI Impacts Skill Formation

AI assistance produces significant productivity gains across professional domains, particularly for novice workers. Yet how this assistance affects the development of skills required to effectively su...

arxiv.org

Meta in effort to fix safety, factually, hallucinations at *pretraining* they ensure the model is trained to generate only high-quality safe tokens, even for unsafe prompts. "Self-Improving Pretraining: using post-trained models to pretrain better models" ( arxiv.org/abs/2601.21343 )

Bild

How will AI agents impact democratic values? Democracies are—for independent reasons—already under acute pressure. Since WWII Moore's Law and democratisation went up and to the right in lockstep. Not any more.

Bild

I spent a few hours with OpenAI's Operator automating expense reports. Most corporate jobs require filing expenses, so Operator could save *millions* of person-hours every year if it gets this right. Some insights on what worked, what broke, and why this matters for the future of agents 🧵

Graph of web tasks along difficulty and severity (cost of errors)

Turns out we weren't done for major LLM releases in 2024 after all... Alibaba's Qwen just released QvQ, a "visual reasoning model" - the same chain-of-thought trick as OpenAI's o1 applied to running a prompt against an image Trying it out is a lot of fun: simonwillison.net/2024/Dec/24/...

Trying out QvQ—Qwen’s new visual reasoning model

I thought we were done for major model releases in 2024, but apparently not: Alibaba’s Qwen team just dropped the Apache2 2 licensed QvQ-72B-Preview, “an experimental research model focusing on …

simonwillison.net