Vikram Saraph

@vikramsaraph.com

Software engineer, AI/ML researcher, and mathematician at Johns Hopkins APL. Former New Englander, current Marylander. Brown CS PhD and Notre Dame math alum. Nerd of sorts (computers, math, language, puzzles, games, books, music). Opinions are my own.

Almost every LLM that reaches a broad audience is quantized—compressed to run cheaply. USC's Emilio Ferrara argues this compression is treated as a safety nonevent, when it should be treated as a change to the deployed system. This has policy implications, he says.

The Model You Audit Is Not the Model You Ship

Almost every LLM that reaches a broad audience is quantized—with under-appreciated ramifications for AI safety, writes Emilio Ferrara.

techpolicy.press

Hmm, not sure I’m a fan of likes of posts by people you follow being more visible now in the UI. I know this data is all public by design anyways, but private likes on Twitter is something I think I prefer.

Nerd thing: so there's this tech/math/nerd culture that's hard to describe, but where there's a kind of enlightened playfulness, which if you like you can trace very far back, and which perhaps hits a kind of high watermark with Claude Shannon, routes through late 60s/70s hacker culture, and so on.