Eugene Vinitsky 🍒

@eugenevinitsky.bsky.social

Anti-cynic. Towards a weirder future. Reinforcement Learning, Autonomous Vehicles, transportation systems, the works. Asst. Prof at NYU. Founding research scientist at Percepta. https://emerge-lab.github.io https://www.admonymous.co/eugenevinitsky

Exciting to announce Discovery Loop II. While many other companies simply focus on automating research, our company focuses on automating the launching of automated research companies. By next year we hope to be launching one of these every minute of the day.

My radical suggestion for peer review: Move from nominal pre-publication review to explicit post-publication review. The original reason for peer review was that journal pages are a limited resource, so we need a filter before publication. That no longer makes sense with digital publishing. 1/

LLMs adopts linguistic quirk -> Humans avoid linguistic quirk -> Humans develop new linguistic quirks -> LLMs adopt linguistic quirk... LLMs can update a lot faster though, and so I wonder if this all does something quite weird to written english.

Are policy gradient methods hopeless without other machinery? Deep learning works well when the Hessian is well conditioned. But the policy objective is under no obligation to give us that. And no amount of architecture engineering can rescue us from problems w/ the objective.

I’m saying this as a routine voluntary user of LLMs: constant questioning and re-evaluation of whether you are interacting with a human does real psychic damage.

An algorithm that makes the people who interact with it better (according to their own stated values) rather than worse is worth pursuing. And, if it turns out to be impossible then it’d be a good idea to shut the whole thing down

"We're all worried," as what it means to do research (in my field, Theoretical CS) seems to be shifting, and shifting fast. What to do? Senior researchers must lead by example, knowing that not everything will pan out. What I'm suggesting below may not work everywhere, but here's my own advice: 1/