Were you watching the World Cup or playing chess? Well, we have the stats for it... (1/2)
t-SNE on SigLIP embeddings of 1,360 @oreilly.bsky.social book cover animals. The penguin cluster is my favorite.
We recently crossed 6 million chess puzzles in our open database. Like the platform itself, Lichess data is free and open source. Go and build something cool with it, Available wherever you get your datasets!
work in progress: 1 million synthetic personas from NVIDIA's Nemotron-Personas-USA dataset.
Been trying to reproduce datashader functionality using only Apache Arrow and Acero as a learning exercise. This is a 1-billion point Clifford attractor rendered in ~8s from a 9GB parquet file. (No JIT)
I'm looking for 5-10 #chess players to test out a tool I'm building. Preferably who play on @lichess.org and are 1200+ in rapid or blitz. And if you coach chess at all, I'd be extra grateful to have you test it! NOTE: it is _NOT_ an "LLM chess coach" tool, I promise! 🙏
Any chess position with 8 pieces on the board and at least one pair of opposed pawns has been solved! Lichess can now tell you definitively if it's a win, loss or draw with no engine required. Read about the massive technological undertaking to accomplish this partial 8 piece tablebase on our blog:
UMAP connectivity plots of 3,627 chess openings from the @lichess.org datasets (huggingface.co/datasets/Lic...)
Which colormap do you think looks the nicest? I'm leaning toward plasma.
Scatterplot of 4 million computer science authors, laid out according to co-authorship connections. Large blob in the bottom left are all single authors; removing them lets the plot breathe more somehow. The source of the data is the @dblp.org bibliography.
Who is winning the open AI race? Our new study Economies of Open Intelligence maps @hf.co 851k models' downloads 2020→2025. 1) Power rebalance: US tech ↓; China + community ↑ 2) Models size & efficient ↑ (MoE, quant, multimodal) 3) Intermediary layers ↑ (adapters/quantizers) 4) Transparency ↓ /🧵
Researchers at Google DeepMind used our free puzzle database and reinforcement learning to train a model to generate creative chess puzzles. ➡️ Read more on this by Tom Zahavy from the DeepMind discovery team: lichess.org/@/tomas135/b...
AI-Generated Chess Puzzles
A new research by the Discovery team at @GoogleDeepMind using RL and generative models to discover creative chess puzzles
lichess.org
Three different ways to represent colo(u)r. Work in progress, inspired by an old post by Kat Zhang / The Poet Engineer.
I made this annotated scatter plot of 1 million FineWeb-Edu documents for @sashamtl.bsky.social's new TED talk.
Also really love how organic the plot looks with "inferno" (left) and "viridis" (right).
Map of the internet: 1.3M nodes (BGP)
Thanks to @jamesabednar.bsky.social I realized I had used the wrong background color for the colormap I had chosen. This is another version of the plot (different embeddings) with the corrected background.
Map of the internet: 1.3M nodes (BGP)
Really cool new embeddings exploration tool by @domoritz.de and colleagues from Apple. Can't wait to build with this. Also includes a streamlit component and a Jupyter widget.
Woah! EA just open sourced "Command and Conquer: Red Alert" and a bunch of other CnC games! github.com/electronicar...
Lichess is now on @kaggle.com! Use our puzzles, openings, and engine evaluation datasets directly in your kaggle notebooks: https://www.kaggle.com/organizations/lichess ♟️
The folks at Foursquare released a @hf.co dataset of 104.5 million places of interest and here's all of them plotted using datashader
I recently used the @lichess.org puzzles dataset to experiment with chess position embeddings and visualize 4.5M starting positions. (hf.co/datasets/Lic...)
The Lichess database of games, puzzles, and engine evaluations is now on @hf.co - https://huggingface.co/Lichess. Billions of chess data points to download, query, and stream and we're excited to see what you'll build with it! ♟️ 🤗