Houjun Liu

@jemoka.com

NLP & POMDPs; CS@Stanford; gradient descent enthusiast www: jemoka.com ac: nlp.stanford.edu/~houjun/

Introducing 𝘁𝗵𝗼𝘂𝗴𝗵𝘁𝗯𝘂𝗯𝗯𝗹𝗲𝘀: a *fully unsupervised* LM for input-adaptive parallel latent reasoning ✅ Learn yourself a reasoning model with normal pretraining ✅ Better perplexity compared to fixed thinking tokens No fancy loss, no chain of thought labels 🚀

Bild

Introducing 𝘁𝗵𝗼𝘂𝗴𝗵𝘁𝗯𝘂𝗯𝗯𝗹𝗲𝘀: a *fully unsupervised* LM for input-adaptive parallel latent reasoning ✅ Learn yourself a reasoning model with normal pretraining ✅ Better perplexity compared to fixed thinking tokens No fancy loss, no chain of thought labels 🚀

Bild

New Paper Day! For EMNLP findings—in LM red-teaming, we show you have to optimize for **both** perplexity and toxicity for high-probability, hard to filter, and natural attacks!

Bild

New Paper Day! For EMNLP findings—in LM red-teaming, we show you have to optimize for **both** perplexity and toxicity for high-probability, hard to filter, and natural attacks!

Bild

I'm excited to announce that I’ll be joining the Computer Science department at Johns Hopkins as an Assistant Professor this Fall! I’ll be working on large language models, computational social science, and AI & society—and will be recruiting PhD students. Apply to work with me!

BildBild

Is it just me or is the latest style guide of ChatGPT's IFT is like terribly sarcastic? I don't need 👉 finger guns after every single message.

I'm to this day still confused about why people are so hyped about softmax-bottlenecked "deep research" approaches; idk about you but I don't usually need to compose a 10 page treatise on first order logic to decide whether or not an if statement is backwards....

cs theory seems like an outlier within theoretical sciences where we are living in the era of Euler, Gauss, etc. for math with substantial results being continuously developed live as a part of frontier research. what a time to be alive.

Ryan Williams@rrwilliams.bsky.social · last yr.

New paper: Simulating Time With Square-Root Space people.csail.mit.edu/rrw/time-vs-... It's still hard for me to believe it myself, but I seem to have shown that TIME[t] is contained in SPACE[sqrt{t log t}]. To appear in STOC. Comments are very welcome!

Ever dreamed of AI agents learning through interacting with the open world unsupervisedly? Our latest preprint introduces NNetNav-Live which collects training data through exploration on real websites and hindsight labeling, which produces a SOTA OSS agent.

Bild

LM agents today primarily aim to automate tasks. Can we turn them into collaborative teammates? 🤖➕👤 Introducing Collaborative Gym (Co-Gym), a framework for enabling & evaluating human-agent collaboration! I now get used to agents proactively seeking confirmations or my deep thinking.(🧵 with video)

Bild

I feel like the instagram onboarding is at this point the primary blocker of me getting an instagram. They *insist* that I’m {fake, spam, using an open proxy} literally no matter what I try or where I try it. So be it, better that way probably too :/ But I do want to post pictoors… maybe Pinterest?