Venkat

@venkatasg.net

Assistant Professor CS @ Ithaca College. Computational Linguist interested in pragmatics & social aspects of communication. venkatasg.net

Maybe this is because of regulations, but its refreshing to see a company (Wise payment transfer) admit that one of their competitors is better in some way (exchange rates in this case) transparently in the UI.

Screenshot of Wise UI showing how much 2000USD is on different services with Wise being second best

Hypo is now out! It provides a rich metadata layer on top of @grain.social records, including gear and workflow abstractions as well as tools for photo annotation and search. If you shoot film, it also provides a stockpile tracker, shot logger, and development timer.

Hypo

Hypo is a tool for organizing and sharing your film or digital photography: the gear you shoot, develop, and scan with; the workflows that take a film roll from capture to finished scan; and the prove...

hypo.graycard.app

Government funded research needs to be better, but this sentence (and paragraph, and whole text) is just gibberish. Lemire says 'stagnant technologically' when he just means that he doesn't find public research in medicine, public policy, social sciences, etc. sexy. But ChatGPT? Now that's sexy 🤦🏾‍♂️

Bild

All the criticisms about LLMs are true. I (and others) use Claude Code/Gemini not because they're useful but because I *like* using them. It's easy to ignore the costs because they're designed to be easy to ignore, and I excuse so much all day already. It's fun to get answers and code instantly.

Nice to see coverage of @wenxuand.bsky.social ‘s work with @gregdnlp.bsky.social and @nickatomlin.bsky.social Paper here: arxiv.org/abs/2602.16699

NSF-Simons AI Institute for Cosmic Origins (CosmicAI)@nsfsimonscosmicai.bsky.social · 2mo ago

Research highlight! CosmicAI Researchers Wenxuan Ding (NYU), @gregdnlp.bsky.social (NYU) as well as external collaborator Nicholas Tomlin (NYU, TTIC) investigated whether LLM agents like Claude Code and OpenAI Codex can navigate cost-benefit tradeoffs in their actions. youtube.com/shorts/GMZ5z...

Spotting the rule from past experience is one thing; acting on it correctly is another. To find out, we introduce HERO's JOURNEY🦸‍♀️ to test for the LLMs’ inductive reasoning ability in multi-step setups. We found models show signs of rule induction, but scratch the surface.😮

BildBild

“Dimicillin” isn’t real. We made it up. Yet many LLMs still call it an antibiotic. Across 9 models and 653 drugs, we find that drug-name affixes alone can drive pharmacological reasoning. Models often rely on morphology over facts. We trace this shortcut from behavior to mechanism. 🧵

BildBild

xkcd #1140 builds ‘Calendar of Meaningful Dates’ from Google Books N-grams corpus, but we have bigger, more diverse corpora now, so what does a Calendar of Meaningful Dates for the Web or a Language Model look like? Using infinigram-mini we can query date forms on DCLM (web text) and The Pile…(1/2)

xkcd #1140 Calendar of meaningful dates from Google n-grams over books

New opinion piece on the interface between research on concepts and categories in minds vs. in neural network LMs! I take the position that there is much to be learned from this interface (e.g., learning about concepts from language alone) and outline some directions for future.

Title page of "Semantic Cognition for and from Language Models" followed by a figure showing tests that target conceptual structure and content vs. those that target function.

What does a scientific figure make you wonder? 📊 We introduce MQUD: multimodal Questions Under Discussion for scientific figures. With 1,250 author-annotated questions over 245 figures from 56 papers, MQUD asks what scientific question a figure raises in context.

Bild

I’m going to try something different and this contest is a good enough playground for this. No agents, LLMs etc. I’m going to (finally) learn JS (and C) properly by trying to port NetHack from C to JS. Going to rely on just textbooks and web search as I used to when learning something new 🤓

DDavid Bau@davidbau.bsky.social · 3mo ago

The Teleport Contest is open. Port NetHack 5.0 from C to JavaScript, bit-exactly. Same screen, every keystroke. Any approach: LLM agents, hand-coded, transpiler, hybrid. Live leaderboard, two phases through December. mazesofmenace.ai/announcement

New paper! 🏁 Last one from my PhD at UT Austin. LLMs sound empathic but repeat the same discourse moves turn after turn — at 2x the rate of humans. We built MINT🌿, the first RL framework for discourse move diversity in empathic dialogue. +25% empathy, −26% repetition. 📄 arxiv.org/abs/2604.11742

Bild

Several replies argued LMs can't be linguistically interesting because they're statistical/ trained on strings. We call this the String Statistics Strawman: language is generative, LMs learn string statistics, since 1957 we know string statistics are insufficient, therefore LMs don't learn language.

table laying out the "String Statistics Strawman" and the "As Good as it gets" assumption in our article.

Announcing a new version of our 2024 paper on linguistic hypothesis generation from LMs! @najoung.bsky.social and I have systematized our hypothesis generation framework, added stringent criteria for model selection, 10x-ed our learning trials, and included an epigraph from Jeff Elman 🙏!

Title page for the paper “A systematic framework for generating novel experimental hypotheses from language models”, with an epigraph from Jeff Elman describing how Rumelhart and McClelland (1982) did hypothesis generation with their connectionist network, and a figure describing our pipeline.

These results suggest new linguistic hypothesis that we are hopeful will be tested in future psycholinguistic work. And more broadly, we believe they add to the growing evidence that LMs have an important role to play in informing linguistic theory! 📃: arxiv.org/abs/2604.13950

Causal Drawbridges: Characterizing Gradient Blocking of Syntactic Islands in Transformer LMs

We show how causal interventions in Transformer models provide insights into English syntax by focusing on a long-standing challenge for syntactic theory: syntactic islands. Extraction from coordinate...

arxiv.org

Perhaps more ink has been spilled on the topic of Island Constraints than any other such phenomenon in theoretical linguistics. In new work with @kmahowald, we study how LMs process such phenomenon, using techniques from mechanistic interpretability to guide our exploration. 🧵👇

Bild