·
AI & ML interests
Contact: arxivgpt@gmail.com
Recent Activity
reacted to theirpost with 🔥 about 6 hours ago 🧬 Your AI can design a malaria drug candidate. Can it tell you whether it's any good?
Open Discovery Challenge #1 — Malaria is live. Design a molecule with any model — OpenAI, Claude, Gemini, Qwen, KIMI, DeepSeek, open weights, or by hand — submit it as SMILES, and it's scored in minutes on whole-cell activity, target binding, selectivity over the human enzyme, ADMET, novelty and synthesisability.
You can check the scoring instead of trusting it. Approved drugs sit on the same leaderboard as the entries: DSM265, a clinical-stage antimalarial, scores 50.9. Teriflunomide — approved, but it hits the human enzyme — scores 2.8. Caffeine scores 1.8. If the clinical candidate lands on top and coffee lands at the bottom, the scorer discriminates.
We caught 14 defects before opening — conventional toxicity cutoffs rejected all three approved antimalarials and coffee. All written up, along with the rule we now hold everything to: a gate that rejects an approved drug is a broken gate.
Your molecule stays yours. No patent interest, nothing into our pipeline. You choose whether it's published — and publishing can cost you patentability, so we say so.
USD 1,000 to the top entry when Season #1 closes 30 September 2026 — not payment for your tokens, but a way of saying the work had worth.
Malaria killed ~597,000 people in 2023, three quarters of them children under five. Not for want of chemistry — for want of a market.
No chemistry needed: the guide ships five prompts you can paste straight into your model, and the full rubric is published.
📖 https://huggingface.co/blog/FINAL-Bench/open-discovery-challenge
🚀 https://huggingface.co/spaces/FINAL-Bench/open-discovery-challenge
Computational assessments of candidates — not measurements, not claims of efficacy. posted an update about 6 hours ago 🧬 Your AI can design a malaria drug candidate. Can it tell you whether it's any good?
Open Discovery Challenge #1 — Malaria is live. Design a molecule with any model — OpenAI, Claude, Gemini, Qwen, KIMI, DeepSeek, open weights, or by hand — submit it as SMILES, and it's scored in minutes on whole-cell activity, target binding, selectivity over the human enzyme, ADMET, novelty and synthesisability.
You can check the scoring instead of trusting it. Approved drugs sit on the same leaderboard as the entries: DSM265, a clinical-stage antimalarial, scores 50.9. Teriflunomide — approved, but it hits the human enzyme — scores 2.8. Caffeine scores 1.8. If the clinical candidate lands on top and coffee lands at the bottom, the scorer discriminates.
We caught 14 defects before opening — conventional toxicity cutoffs rejected all three approved antimalarials and coffee. All written up, along with the rule we now hold everything to: a gate that rejects an approved drug is a broken gate.
Your molecule stays yours. No patent interest, nothing into our pipeline. You choose whether it's published — and publishing can cost you patentability, so we say so.
USD 1,000 to the top entry when Season #1 closes 30 September 2026 — not payment for your tokens, but a way of saying the work had worth.
Malaria killed ~597,000 people in 2023, three quarters of them children under five. Not for want of chemistry — for want of a market.
No chemistry needed: the guide ships five prompts you can paste straight into your model, and the full rubric is published.
📖 https://huggingface.co/blog/FINAL-Bench/open-discovery-challenge
🚀 https://huggingface.co/spaces/FINAL-Bench/open-discovery-challenge
Computational assessments of candidates — not measurements, not claims of efficacy. View all activity Organizations
published an article about 7 hours ago view article Who Tells You Whether the Molecule Your AI Just Designed Is Any Good?
FINAL-Bench
• • 9
view article AX-Ray, Finding Causal-Leakage Defects in Two General-Purpose Public Models
FINAL-Bench
• • 13
view article The Fast Gemma Challenge: our verified-SOTA recipe, in full
FINAL-Bench
• • 24
view article POCKET: a 35-billion-parameter model that runs on your iPhone — and on your PC with no GPU
FINAL-Bench
• • 12
view article Aether-7B-5Attn: A 100% Open-Source Sovereign Foundation Model — and a Controlled Experiment in Heterogeneous Attention
FINAL-Bench
• • 21
published an article about 1 month ago view article VKUE: No GPU? Runs Anyway — a 34.7B Reasoner on a Laptop and on Bare CPU
FINAL-Bench
• • 17
published an article about 1 month ago view article Quantum Cryptanalysis on Real Hardware: Pushing Symmetric-Structure Key Recovery Beyond the Published Frontier
published an article about 1 month ago published an article about 2 months ago view article Chitos: From Detection to Proof — An Autonomous Security AI That Actually Exploits
FINAL-Bench
• • 19
view article FINAL-Bench Quantum: An Open, Neutral Benchmark for Quantum-Computing Methods
FINAL-Bench
• • 17
view article Training-Free Reasoning at 88.89% on GPQA Diamond: How Darwin Family Hit Frontier Scores Without a Single Gradient Step
FINAL-Bench
• • 18
view article Darwin-TTS: We Gave a TTS Model 3% of an LLM's Brain — It Started Showing Emotion
FINAL-Bench
• • 13
view article "Darwin-27B-Opus: Surpassing the Foundation Model Without Training"
FINAL-Bench
• • 16
view article Darwin V6: Diagnostic-Guided Evolutionary Model Merging
view article "The Child That Surpassed Both Parents Through MRI-Guided Evolutionary Merge"
FINAL-Bench
• • 15
view article Introducing WM Bench: A Benchmark for Cognitive Intelligence in World Models
FINAL-Bench
• • 13
view article 🏟️ Smol AI WorldCup: A 5-Axis Benchmark That Reveals What Small Language Models Can Really Do
FINAL-Bench
• • 38
view article MARL: Runtime Middleware That Reduces LLM Hallucination Without Fine-Tuning
view article Structural Problems in AI Benchmarking and the Case for a Unified Evaluation Framework
view article Do Bubbles Form When Tens of Thousands of AIs Simulate Capitalism?
FINAL-Bench
• • 17