·
AI & ML interests
Contact: arxivgpt@gmail.com
Recent Activity
Organizations
published an article 43 minutes ago view article Did the Civilization Emerge, or Was It Recited?
FINAL-Bench
• • 1
view article Writing Down the Line Between Luck and Skill
FINAL-Bench
• • 14
view article We changed one line and the benchmark score moved 0.21 AUROC
FINAL-Bench
• • 13
view article Who Tells You Whether the Molecule Your AI Just Designed Is Any Good?
FINAL-Bench
• • 11
view article AX-Ray, Finding Causal-Leakage Defects in Two General-Purpose Public Models
FINAL-Bench
• • 13
view article The Fast Gemma Challenge: our verified-SOTA recipe, in full
FINAL-Bench
• • 24
published an article about 1 month ago view article POCKET: a 35-billion-parameter model that runs on your iPhone — and on your PC with no GPU
FINAL-Bench
• • 12
published an article about 1 month ago view article Aether-7B-5Attn: A 100% Open-Source Sovereign Foundation Model — and a Controlled Experiment in Heterogeneous Attention
FINAL-Bench
• • 21
published an article about 2 months ago view article VKUE: No GPU? Runs Anyway — a 34.7B Reasoner on a Laptop and on Bare CPU
FINAL-Bench
• • 17
published an article about 2 months ago view article Quantum Cryptanalysis on Real Hardware: Pushing Symmetric-Structure Key Recovery Beyond the Published Frontier
published an article about 2 months ago view article Chitos: From Detection to Proof — An Autonomous Security AI That Actually Exploits
FINAL-Bench
• • 19
view article FINAL-Bench Quantum: An Open, Neutral Benchmark for Quantum-Computing Methods
FINAL-Bench
• • 17
view article Training-Free Reasoning at 88.89% on GPQA Diamond: How Darwin Family Hit Frontier Scores Without a Single Gradient Step
FINAL-Bench
• • 18
view article Darwin-TTS: We Gave a TTS Model 3% of an LLM's Brain — It Started Showing Emotion
FINAL-Bench
• • 13
view article "Darwin-27B-Opus: Surpassing the Foundation Model Without Training"
FINAL-Bench
• • 16
view article Darwin V6: Diagnostic-Guided Evolutionary Model Merging
view article "The Child That Surpassed Both Parents Through MRI-Guided Evolutionary Merge"
FINAL-Bench
• • 15
view article Introducing WM Bench: A Benchmark for Cognitive Intelligence in World Models
FINAL-Bench
• • 13
view article 🏟️ Smol AI WorldCup: A 5-Axis Benchmark That Reveals What Small Language Models Can Really Do
FINAL-Bench
• • 38