lea

@leeps.bsky.social

reposting art, research https://www.leandra.dev/

Going through some Erdős problems that Mark Sellke and Mehtaab Sawhney found had already been solved using GPT5 pro for search. More precisely, these were listed as open, but solutions to them already existed in the literature, but not recognized as such. Quite remarkable. One example: [+]

Bild

🌌🛰️🔭Want to explore universal visual features? Check out our interactive demo of concepts learned from our #ICML2025 paper "Universal Sparse Autoencoders: Interpretable Cross-Model Concept Alignment". Come see our poster at 4pm on Tuesday in East Exhibition hall A-B, E-1208!

Harry Thasarathan@hthasarathan.bsky.social · last yr.

🌌🛰️🔭Wanna know which features are universal vs unique in your models and how to find them? Excited to share our preprint: "Universal Sparse Autoencoders: Interpretable Cross-Model Concept Alignment"! arxiv.org/abs/2502.03714 (1/9)

Harmony merges vision-language training and self-supervised learning, achieving strong results on dense tasks with web-scraped data. It outperforms CLIP and SLIP, showing better performance in zero-shot classification and segmentation across various benchmarks. https://arxiv.org/abs/2405.14239

Harmony: A Joint Self-Supervised and Weakly-Supervised Framework for Learning General Purpose Visual Representations

ArXiv link for Harmony: A Joint Self-Supervised and Weakly-Supervised Framework for Learning General Purpose Visual Representations

arxiv.org

a thread of the most exciting things I saw on #CVPR2025 day 1! (1) TerraMesh from IBM + ESA It’s a massive, multimodal dataset for Earth foundation models. Think ImageNet, but for the planet. 9M+ globally aligned image patches combining radar, optical, elevation, vegetation, and land use. 👇

Bild

Not Gaussian or NeRF! SuperDec rocks 3D scenes with superquadrics—compact, clever, and oh-so-smooth! This method uses transformer-based nets to predict superquadric params and soft segmentation, then refines them with Levenberg–Marquardt optimization.

We propose Neurosymbolic Diffusion Models! We find diffusion is especially compelling for neurosymbolic approaches, combining powerful multimodal understanding with symbolic reasoning 🚀 Read more 👇

this workshop combined my loves: an abandoned wework + open source AI we built AI agents with open-source tools (smolagents, Qwen), and then used them (and our in-house robot arm) to generate and assemble mocktails! 🤖🍹 Huge thanks to 🤗 @hf.co for sponsoring & for creating smolagents 🦾

workshop @ frontier tower

Happy birthday Arne Wolf. The under-the-radar artist, calligrapher, & educator, remembered as an influential teacher, experimental typographer, & bookmaker, was born today in 1929. “The transition from craft to art is equal to the movement from making letters to making content."

Bild

New Paper: Continuous Thought Machines pub.sakana.ai/ctm/ Neurons in brains use timing and synchronization in the way that they compute, but this is largely ignored in modern neural nets. We believe neural timing is key for the flexibility and adaptability of biological intelligence. Thread ↓

Sakana AI@sakanaai.bsky.social · last yr.

“Continuous Thought Machines” Blog → sakana.ai/ctm Modern AI is powerful, but it's still distinct from human-like flexible intelligence. We believe neural timing is key. Our Continuous Thought Machine is built from the ground up to use neural dynamics as a powerful representation for intelligence.

C5: Web-Scale Creative Commons Web Data - Effort to collect CC-licensed web data in one place - 147M documents explicitly CC-licensed - License detection from HTML links & metadata - Preserves exact license type & version - Maps to original URLS & crawl dates huggingface.co/datasets/Bra...

BramVanroy/CommonCrawl-CreativeCommons · Datasets at Hugging Face

We’re on a journey to advance and democratize artificial intelligence through open source and open science.

huggingface.co