2025
view article2023
view article2023
view article2022
view article2017
view article2016
view article2005
view article2000
view article1994
view article1989
view article1988
view article1986
view article1978
view article1975
view article1974
view article1973
view article1971
view article1970
view article1966
view article1962
view article1960
view article1952
view article1948
view article1945
view article1943
view article1935
view article1934
view article1925
view article1923
view article1921
view article1916
view article1914
view article1905
view article1900
view article1898
view article1888
view article1888
view article1883
view article1863
view article1860
view article1855
view article1831
view article1819
view article1813
view article1810
view article1796
view article1781
view article1761
view article1760
view article1727
view article1726
view article1655
view article1565
view article1522
view article1514
view article1380
view article1334
view article1276
view article1264
view article1253
view article1198
view article617
view article70
view article14
view articleIf you're giving me e-paper, I'm going to make an e-printer.
I wrote recently about how the collection of good, fruitful open problems is now being mined in a non-renewable fashion, leading to the potential scenario of these problems becoming scarce. This may seem unintuitive at first, since the set of possible problems one could ask is infinite. Perhaps the following analogy can help: a country or region can suffer a critical shortage of drinking water while simultaneously being surrounded by a massive ocean. One can easily generate any number of open problems in mathematics at will, such as working out the 10^10^10th digit of pi. But the vast majority of such problems are not worth focusing attention on: they show no particular propensity to reveal any further insights or connections to other q
A visualization of the attention mechanism in LLMs.
The notes for this blog post have been sitting in my drafts folder for half a year now. I’ve done a little work on NumPy itself in the past year. Nothing…
Mercury 2.5 is the most capable diffusion LLM on the market. It runs at 1,107 tokens/sec and offers a 40% increase in intelligence over Mercury 2, comparable to cost-optimized frontier models.
I test Unsloth GGUFs of Qwen3.8 27B (Q4_K_M, UD-Q2_K_XL, UD-IQ1_S) with llama.cpp on GPQA Diamond, IFBench, Terminal-Bench 2.1. Q4_K_M matches BF16 abd fits an RTX 4090.
Sometimes you find gold in unexpected places. In this case, it's from a video about the CH-53 helicopter. The CH-53 Sea Stallion first entered service in 1966. It's a huge helicopter, capable of lifti...
We’re sharing an AI-generated solution to the Navier–Stokes Millennium Prize Problem, including a writeup and a formal proof in Lean.
A skill to stop your coding agent from burying the answer. ADHD-friendly output. - ayghri/i-have-adhd
Meet Muse, Meta's personal AI agent. Learn what it can do across everyday tasks, how it works, and how it helps you get more done.
ARGODRIVE Deltafin: Kimi K3 (2.8T MoE) streamed from SSDs on Apple Silicon — fork of gavamedia/deltafin with the ARGODRIVE storage work and benchmark package - argonautlabsai/deltafin
We’re introducing AlphaGenome Atlas, a database predicting the effects of every possible single nucleotide variant in the human genome.