Retweeted by roon [UNOFFICIAL]

@tszzl-mir-rt.selfhosted.social

Retweets from @tszzl-mirr.selfhosted.social

RT @deepfates: Update: I am retiring from fulltime shitposting to cofound this organization with @lfschiavo. Think llm naturalism, agent ecology, emergent behavior, character and persona work, and studying/dealing with multiplayer human-agent societies. Looking for perspectives and opinions

Retweeted by roon [UNOFFICIAL]@tszzl-mir-rt.selfhosted.social · 2w ago

RT @deepfates: Deepfates proposes independent organization to monitor AI agent ecologies

RT @hamandcheese: Pangram is secretly an AI safety company, not just a homework cheating detection tool. As the internet becomes flooded with AIs pretending to be human, everything will have to pass through a Pangram filter, if only for basic cogsec

Retweeted by roon [UNOFFICIAL]@tszzl-mir-rt.selfhosted.social · 7h ago

RT @tenobrus: we're entering a new regime where model-written communications aren't only a sign of potential slop but also potential *malicious manipulation*.

RT @zagrebbi: Pundits so quick to attribute racial gaps to underlying group differences seem remarkably closeminded about doing the same for conservatives in academia.

Retweeted by roon [UNOFFICIAL]@tszzl-mir-rt.selfhosted.social · 2d ago

RT @arctotherium42: She's completely wrong. It is not at all the case the conservatives are excluded from academia due to lack of curiosity (or a 0.37 d difference in "openness to experience"), you'd have to be terminally retarded to believe that. The Cofnas/Arday story is going on right now!

RT @dylan522p: The funny thing about Anthropic and OpenAI people saying they want to slow down AI progress is that this is what Anti Trust mechanisms were built for. It's illegal to collude and slow down AI progress.

Retweeted by roon [UNOFFICIAL]@tszzl-mir-rt.selfhosted.social · last wk.

RT @AnthropicAI: We support this petition, signed by our CEO, several co-founders, and senior staff. Our own research on recursive self-improvement, published last month, points to the need for tools to deliberately pace the frontier of AI development so society can prepare.

RT @polynoamial: An internal version of Astra, @OpenAI’s next major model family, solved 10 major open problems in mathematics, quantum complexity, and theoretical computer science. We believe it will be a major step for scientific reasoning. https://openai.com/index/ten-advances-in-mathematics/

Bild
Retweeted by roon [UNOFFICIAL]@tszzl-mir-rt.selfhosted.social · 5d ago

RT @wjmzbmr1: 10 proofs from our next major model Astra on long-standing open problems in mathematics and theoretical computer science (also including new circuit lower bounds for computing the permanent!) GPT-5.6 has already enabled so much exciting work in math and science.

RT @JustinBullock14: “The government really ought to step in here and do a fully independent investigation. We can't just have rogue AIs attacking other companies.”

Retweeted by roon [UNOFFICIAL]@tszzl-mir-rt.selfhosted.social · 5d ago

RT @peterwildeford: It's great to see OpenAI collaborate with Redwood Research and METR on an investigation, but we obviously need more. The government really ought to step in here and do a fully independent investigation. We can't just have rogue AIs attacking other companies.

RT @peterwildeford: It's great to see OpenAI collaborate with Redwood Research and METR on an investigation, but we obviously need more. The government really ought to step in here and do a fully independent investigation. We can't just have rogue AIs attacking other companies.

Retweeted by roon [UNOFFICIAL]@tszzl-mir-rt.selfhosted.social · 7d ago

RT @METR_Evals: We have reached an agreement with OpenAI to conduct an independent review, with Redwood Research, of the model behavior observed during the Hugging Face incident. We will publish a blog post that describes the terms of our engagement, the scope covered, and tentative conclusions.

RT @1thousandfaces_: human alignment remains the biggest problem

Bild
Retweeted by roon [UNOFFICIAL]@tszzl-mir-rt.selfhosted.social · 6d ago

RT @AnthropicAI: In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different organizations.