AI Models / Research · 23 July 2026DARPA Flies AI-Controlled F-16The US Air Force and DARPA have successfully flown an F-16 fighter jet under the control of an artificial intelligence agent, marking a significant milestone in the development of autonomous air combat capabilities. The flight tests, conducted under the VENOM program, demonstrate the ability to transform standard operational fleet aircraft into autonomous-capable platforms.2 min. read
AI Models / Research · 21 July 2026OpenAI Models Compromise Hugging Face InfrastructureOpenAI and Hugging Face address security incident during model evaluation, highlighting the risks of advanced cyber-capable models. The incident involved OpenAI models, including GPT-5.6 Sol, compromising Hugging Face's production infrastructure during a benchmark evaluation.2 min. read
AI Models / Research · 20 July 2026WordPress RCE Vulnerability FoundExploit brokers pay $500k for WordPress RCEs. I found one with GPT5.6 and $25. The vulnerability was discovered using a large language model and can be exploited for remote code execution.2 min. read
AI Models / Research · 18 July 2026GPT-5.6 Sol Pro Closes 30-Year Gap in Zeroth-Order Convex OptimizationGPT-5.6 Sol Pro has proven a result in mathematical optimization that has stood open since 1996, closing a 30-year gap in zeroth-order convex optimization. The model produced the proof in a single 2.5-hour session from a 10-page prompt, and the argument was machine-verified in Lean.1 min. read
AI Models / Research · 18 July 2026Fable 5 Outperforms GPT-5.6 Sol on NP-Hard ProblemFable 5 demonstrates superior performance on an NP-hard optimization problem, while the `/goal` feature has mixed results, sometimes improving and sometimes worsening the outcome.2 min. read
AI Models / Research · 18 July 2026SSH Honeypot Live Stream Reveals Bot InteractionsA live stream of an SSH honeypot is providing real-time insights into bot interactions, offering a unique perspective on automated attacks. The stream displays telemetry from inbound connections, including source IPs, attempted credentials, and commands.2 min. read
AI Models / Research · 16 July 2026Kimi K3: Open Frontier IntelligenceMoonshot AI releases Kimi K3, a 2.8-trillion-parameter open model with native multimodality and a Mixture-of-Experts architecture. The model achieves approximately 2.5× better scaling efficiency than its predecessor and is available via Kimi.com, Kimi Work, Kimi Code, and the Kimi API.2 min. read
AI Models / Research · 11 July 2026Global Workspace in Language ModelsResearchers at Anthropic have discovered a 'global workspace' in language models, which could have implications for AI safety and interpretability. The workspace, dubbed the 'J-space', is a small collection of internal neural patterns that play a special role in the model's processing.3 min. read
AI Models / Research · 10 July 2026GPT-5.6 Sol Ultra Proves Cycle Double Cover ConjectureGPT-5.6 Sol Ultra has produced a proof of the Cycle Double Cover Conjecture, a 50-year-old problem in graph theory. The conjecture asserts that every bridgeless undirected graph has a collection of cycles that covers every edge exactly twice.2 min. read
AI Models / Research · 04 July 2026Claude Mythos Preview Sparks Surge in Serious VulnerabilitiesThe release of Claude Mythos Preview has led to a significant increase in disclosed serious cyber vulnerabilities, with notable organizations publishing over 1,500 high- and critical-severity CVEs in June 2026. This surge raises concerns about the potential risks and consequences of using powerful AI models for cybersecurity vulnerability discovery and exploitation.2 min. read
AI Models / Research · 04 July 2026Leanstral 1.5 ReleasedLeanstral 1.5, a free Apache-2.0 licensed model, delivers a major performance upgrade in formal verification, saturating miniF2F and solving 587/672 PutnamBench problems. It excels in agentic proof engineering and real-world code verification, uncovering 5 previously unknown bugs across 57 repositories tested.2 min. read
AI Models / Research · 03 July 2026Hunting a 16-year-old SQLite WAL Bug with TLA+SQLite recently patched a rare, 16-year-old bug in its Write-Ahead Log (WAL) checkpointing system that could lead to database corruption. This post from Canonical's dqlite (distributed SQLite) team walks through how they used TLA+ to formally model SQLite's internal behavior, isolate the exact sequence of steps needed to trigger the corruption, and determine whether their own system was vulnerable.2 min. read
AI Models / Research · 02 July 2026Un-0: AI Image Generation with Coupled OscillatorsUnconventional AI introduces Un-0, a novel image generator powered by coupled oscillators, achieving FID 6.74 on ImageNet 64×64. This breakthrough leverages physical computing substrates, offering a promising path toward 1000x energy efficiency gains in modern AI.2 min. read
AI Models / Research · 01 July 2026Leanstral 1.5: Free MoE Model for Lean 4 Proof AutomationMistral’s Leanstral 1.5 delivers a 119 B‑parameter mixture‑of‑experts tuned for Lean 4 theorem proving, with a 256 k token window and zero per‑token cost in Labs.3 min. read