Open-Source AI Finally Caught Up?
20th July 2026 | Superintelligence Newsletter
Hey Superintelligence Fam 👋
Open models are no longer chasing from behind. Kimi K3, Inkling, and Hy3 show that scale, reasoning, multimodality, and enterprise readiness are rapidly becoming shared advantages across the AI landscape.
This week also reveals AI’s next battleground: not just smarter models, but better agents, stronger metacognition, meaningful routing, responsible governance, and new interfaces that make intelligence easier to use everywhere.
Let’s dive into what’s new this week..
Become AI Product Management Certified | September Cohort
Join 6-week live, mentor-led cohort designed for both tech and non-tech professionals. You will learn product management fundamentals, identify valuable AI opportunities, design responsible AI solutions, validate real product ideas, and build a working AI product for your portfolio. September cohort registrations are now open, with limited seats available.
Kimi K3 Launches With 2.8 Trillion Parameters : Kimi K3 arrives with 2.8 trillion parameters, a 1-million-token context window, native vision, always-on reasoning, and 2.5× K2 scaling efficiency for coding and knowledge work.
Thinking Machines Debuts Open-Weight Inkling : Thinking Machines debuts Inkling, a 975-billion-parameter open-weight multimodal model activating 41 billion parameters per task, trained on 45 trillion tokens and built for enterprise customization.
Apple Lawsuit Escalates Pressure on OpenAI : Apple’s lawsuit targets OpenAI’s hardware ambitions after more than 400 Apple hires and a $6.5 billion acquisition, while employee activism and regulatory backlash intensify.
OpenAI Launches a $230 Codex Command Center : OpenAI and Work Louder unveil the $230 Codex Micro, featuring 13 mechanical switches, live RGB agent status, a reasoning dial, joystick controls, and 32 custom keycaps.
Kimi K3 : Kimi K3 packs 2.8 trillion parameters, activates 16 of 896 experts, handles 1M-token contexts, supports vision, and delivers roughly 2.5× K2 scaling efficiency for agents.
Grok 4.5 : Grok 4.5 runs at 80 tokens per second, scores 83.3% on Terminal-Bench 2.1, uses 4.2× fewer tokens, and costs $2/$6 per million tokens via API.
Tencent Hy3 : Tencent’s open-weight Hy3 combines 295B parameters, 21B active parameters, 256K context, 192 experts, and cuts hallucinations from 12.5% to 5.4% after feedback from 50-plus products.
Gemini Notebook : Rebranded from NotebookLM, Gemini Notebook turns PDFs, websites, YouTube, audio, Docs, and Slides into grounded answers, Audio Overviews, videos, infographics, presentations, and research connections quickly.
Self-Improvements in Modern Agentic Systems: A Survey : This survey unifies self-improving agents as foundation models plus scaffolds, systematically mapping updates across weights, prompts, memory, tools, control logic, applications, evaluation, and open challenges.
Metacognition in LLMs: Foundations, Progress, and Opportunities : Yale and UC Irvine’s survey frames metacognition as monitoring plus control, organizing confidence, calibration, self-verification, abstention, hallucination reduction, and knowledge-boundary detection into one research agenda.
When Is Routing Meaningful? Diversity and Robustness in Language Model Societies : DeepMind researchers show routing needs diversity and paraphrase stability, finding curated societies of fewer than ten agents recover most diversity while KNN routers remain fragile.
Rethinking the Evaluation of Harness Evolution for Agents : On Terminal-Bench 2.1, harness evolution averaged 67.4, trailing the 68.2 baseline, 72.3 parallel sampling, and 71.8 harness scaling, challenging claimed self-improvement gains under matched compute budgets.
Australia announced an Office of AI and environmental rules for data centres; DeepMind proposed a global frontier-model watchdog; Indonesia drafted AI copyright protections covering disclosure, creator compensation and style imitation; while the UK advanced child-safety restrictions targeting manipulative chatbots, AI bias and harmful synthetic content.
Become AI Product Management Certified | September Cohort
Join 6-week live, mentor-led cohort designed for both tech and non-tech professionals. You will learn product management fundamentals, identify valuable AI opportunities, design responsible AI solutions, validate real product ideas, and build a working AI product for your portfolio. September cohort registrations are now open, with limited seats available.
Thank you for tuning in to this week’s edition of Superintelligence Newsletter! Stay connected for more groundbreaking insights and updates on the latest in AI and superintelligence.
For more in-depth articles and expert perspectives, visit our website | Have feedback? Provide feedback.
To explore Superintelligence Media : Explore Here
Stay curious, stay informed, and keep pushing the boundaries of what’s possible!
Until Next Time!
Superintelligence Team.









