Edward Grefenstette

@egrefen.bsky.social

FR/US/GB AI/ML Person, Director of Research at Google DeepMind, Honorary Professor at UCL DARK, ELLIS Fellow. Ex Oxford CS, Meta AI, Cohere.

Do you have a PhD (or equivalent) or will have one in the coming months (i.e. 2-3 months away from graduating)? Do you want to help build open-ended agents that help humans do humans things better, rather than replace them? We're hiring 1-2 Research Scientists! Check the 🧵👇

FYI this posting for a research scientist position in the autonomous assistants team at Google DeepMind will be open for a little under a week, as of today. Please consider applying if you are interested and qualify. See post for details, or ask questions here.

Edward Grefenstette@egrefen.bsky.social · last yr.

Our team in London is hiring a research scientist! If you want to come work with a wonderful group of researchers on investigating the frontiers of autonomous open-ended agents that help humans be better at doing things we love, come have a look. Link in post below 👇

Our team in London is hiring a research scientist! If you want to come work with a wonderful group of researchers on investigating the frontiers of autonomous open-ended agents that help humans be better at doing things we love, come have a look. Link in post below 👇

Researchers: be constructively skeptical about LLMs. Find where they don't work by building with them. Find out if the failure is systemic or just transient. This way, you're best positioned to build what's next, or, if they keep working, to benefit from their growth.

Seek novelty in what you do, how you do it, and who you do it with. I feel part of happiness lies in committing to these things, but not obsessively overcommitting to just one of these things.

Multi-agent peeps: are there any *-MDP variants where there is more than one agent, but exactly one agent is acting on the environment at each time step? Not in the sense of "we take turns" (although I guess it's a special case) but more in the sense that the agents decide who gets to act...

🚨 LLMs can learn to reason from procedural knowledge in pretraining data! 🚨 I particularly enjoy research where the evidence contradicts our initial hypothesis. If you're interested in LLM reasoning, check out the 60+ pages of in-depth work at arxiv.org/abs/2411.12580

Laura@lauraruis.bsky.social · 2y ago

How do LLMs learn to reason from data? Are they ~retrieving the answers from parametric knowledge🦜? In our new preprint, we look at the pretraining data and find evidence against this: Procedural knowledge in pretraining drives LLM reasoning ⚙️🔢 🧵⬇️

“LLMs can/can’t reason” — whatever you think, they clearly can solve some reasoning problems, but how do they learn to do this? Is the dependency on the training data measurable, relative to factual knowledge? Does this tell us something about their abilities? Find out here!

Laura@lauraruis.bsky.social · 2y ago

How do LLMs learn to reason from data? Are they ~retrieving the answers from parametric knowledge🦜? In our new preprint, we look at the pretraining data and find evidence against this: Procedural knowledge in pretraining drives LLM reasoning ⚙️🔢 🧵⬇️

How do LLMs learn to reason from data? Are they ~retrieving the answers from parametric knowledge🦜? In our new preprint, we look at the pretraining data and find evidence against this: Procedural knowledge in pretraining drives LLM reasoning ⚙️🔢 🧵⬇️

🌶️(?) take: Agents are somehow hot right because people realized that LLM output can be interpreted as a DSL which directs side effects in the world (e.g. tool calls) rather than just returning text in a chat/autocomplete sense. What are the open challenges? A 🧵... [1/11]

Is there some good way to selectively crosspost to X and Bluesky, e.g. draft a post somewhere central, and then just post to one/the other/both with a keypress or click? Obviously I can just copy/paste... maybe that's the easiest way.