ai-trend-notifier

$ tree wiki/

Wiki

The compounding AI-trend knowledge base: organizations, models, concepts, people, and papers in one place.

107 pages

Entities

/15
Alibaba / Qwen AI LabHangzhou-based Chinese technology conglomerate. In AI: principally known for the Qwen family of open-weight l…updated 2026-07-2012 in-linksAnthropicSan Francisco-based AI safety company. Develops the frontier claude model series. Strongly positioned in AI s…36 in-linksAppleConsumer electronics and software company. Not a frontier AI lab, but as of WWDC 2026 (June 8), Apple has rep…updated 2026-07-125 in-linksDeepSeekChinese AI research lab (affiliated with High-Flyer Capital Management, Hangzhou). Known for releasing fronti…updated 2026-07-252 in-linksGoogle DeepMindGoogle's AI research division (the 2023 merger of DeepMind + Google AI). Develops the Gemini model series. Fo…39 in-linksMeituanChinese consumer technology conglomerate (food delivery, local services, travel). In 2026, Meituan's internal…updated 2026-07-053 in-linksMeta AIMeta's AI research division. FAIR (Fundamental AI Research) plus the recently established Meta Superintellige…11 in-linksMicrosoftTech giant based in Redmond, WA. Dominates the AI developer and enterprise tooling market through GitHub Copi…10 in-linksMiniMaxShanghai-based AI startup (founded 2021). Develops frontier multimodal models and consumer AI products. Best …updated 2026-07-213 in-linksMistral AIParis-based AI research and product company. Open-weight model lineup (Mistral, Mixtral) + Le Chat chatbot + …9 in-linksMoonshot AIBeijing-based AI startup, creators of the Kimi assistant and model family. Known for pushing long-context cap…updated 2026-07-203 in-linksNVIDIAA GPU manufacturer and the de facto standard for AI compute infrastructure. The infrastructure core of the AI…10 in-linksOpenAISan Francisco-based frontier AI lab. Developer of the GPT series (ChatGPT). 2026 slogan: "Year of Science". F…33 in-linksxAIAI research company founded by Elon Musk (2023). Develops the Grok model series. Leverages real-time data acc…14 in-linksZ.aiZ.ai is the international brand of Zhipu AI (智谱AI), a Beijing-based AI company founded in 2019, spun out of T…updated 2026-07-136 in-links

Models

/51
AlphaEvolveDesigning advanced algorithms via evolutionary searchupdated 2026-07-217 in-linksClaude Fable 5June 9, 2026 — the first publicly available Mythos-class model from Anthropic. Fable 5 and Claude Mythos 5 sh…updated 2026-07-2117 in-linksClaude Mythos PreviewAnnounced April 7, 2026 alongside Project Glasswing. Withheld from commercial release. Anthropic explicitly s…updated 2026-07-219 in-linksClaude Opus 4.7Thoroughness / consistencyupdated 2026-07-2110 in-linksClaude Opus 4.82026-05-28 — released alongside the Anthropic $65B Series H funding announcement.updated 2026-07-2112 in-linksClaude Opus 5Max output (sync): 128k tokensupdated 2026-07-252 in-linksClaude ScienceClaude Science is a scientific research workbench that gives researchers a unified environment for computatio…updated 2026-07-212 in-linksClaude Sonnet 5June 30, 2026 — released the same day as Fable 5/Mythos 5 export-control restoration and Claude Science launc…updated 2026-07-215 in-linksCo-Scientist (Google DeepMind)Research demo / early access: 2026-02 (initial Co-Scientist blog)updated 2026-07-213 in-linksCosmos 3 Super2026-06-01 (Cosmos 3 series launched together: Super / Nano / Edge pending)updated 2026-07-214 in-linksDeep Research MaxAnnounced April 22, 2026 alongside Deep Research (the speed-optimized tier) and the Gemini Enterprise Agent P…updated 2026-07-212 in-linksDeepSeek V4Preview: April 24, 2026. GA: mid-July 2026.updated 2026-07-202 in-linksDevstral 2May 2026 (exact date TBD — flagged in sources as a May 2026 release)updated 2026-07-215 in-linksGemini 3.1 Deep ThinkAdvanced "Deep Think" version: ~January 2026updated 2026-07-217 in-linksGemini 3.5 FlashOutperforms Gemini 3.1 Pro across the full coding and agentic benchmark suite.updated 2026-07-216 in-linksGemini 3.5 Flash CyberChrome V8 JavaScript engine vulnerability finding:updated 2026-07-224 in-linksGemini 3.5 Flash-LiteHigh-throughput, low-latency agentic pipelines (agentic search, document processing)updated 2026-07-224 in-linksGemini 3.5 ProAnnounced at Google I/O 2026 (May 19, 2026) with a June 2026 general-availability target. As of July 17, 2026…updated 2026-07-228 in-linksGemini 3.6 FlashKnowledge cutoff: March 2026 (up from January 2025 on 3.5 Flash).updated 2026-07-224 in-linksGemini OmniGemini Omni represents "a leap forward in world understanding, multimodality and editing" — Google's framing …updated 2026-07-211 in-linksGemini Robotics ER 1.6Released April 15, 2026. Successor to Gemini Robotics-ER 1.5. Notable collaboration with Boston Dynamics on i…updated 2026-07-215 in-linksGemini SparkGemini Spark is a 24/7 cloud-based personal agent that takes actions on behalf of users even when they're off…updated 2026-07-212 in-linksGemma 3nEarly preview released 2026-05-12.updated 2026-07-211 in-linksGemma 4 12BFull precision: 16GB VRAM (RTX 4060, RTX 5090, Apple Silicon Mac)updated 2026-07-212 in-linksGLM-5.2June 13, 2026: Available to Z.ai GLM Coding Plan subscribersupdated 2026-07-215 in-linksGPT-5.5 InstantA claim of simultaneous improvement across three axes (intelligence, clarity, personalization).updated 2026-07-213 in-linksGPT-5.6 Sol (and Terra, Luna)GPT-5.6 is OpenAI's three-tier model family, announced June 26, 2026:updated 2026-07-215 in-linksGPT-Live-1Full-duplex speech: model listens and speaks at the same time; users can interrupt naturally mid-sentenceupdated 2026-07-212 in-linksGPT-Realtime-2 (OpenAI)2026-05-07. Simultaneously: Realtime API exits beta → generally available for production.updated 2026-07-213 in-linksGPT-Rosalind2026-04-16 (model release); 2026-05-29 (Biodefense program launch)updated 2026-07-215 in-linksGrok 4.1 Fast (xAI)May 2026 (exact date unconfirmed; live on x.ai/api as of May 2026)updated 2026-07-213 in-linksGrok 4.5xAI's enterprise deployment of the V9-class model. 1.5 trillion parameters, Cursor-trained — entered private …updated 2026-07-214 in-linksGrok 4.6Announced: July 18, 2026 (Elon Musk on X)updated 2026-07-231 in-linksGrok Build2026-06-22: /goal mode — long-running autonomous execution (plan→execute→verify) for SuperGrok/X Premium+updated 2026-07-217 in-linksGrok Imagine Video 1.5 (Preview)2026-06-03 (API preview)updated 2026-07-213 in-linksGrok V9-MediumxAI's coding-focused foundation model — 1.5-trillion parameters, ~3× larger than the prior production Grok. T…updated 2026-07-214 in-linksKimi K3Frontend Code Arena: beats Anthropic Fable 5 (human-preference Elo) — Moonshot's reported benchmarkupdated 2026-07-204 in-linksLeanstral 1.5Formal verification: generating Lean 4 proofs for functions and algorithmsupdated 2026-07-213 in-linksLongCat-2.0LongCat-2.0 had been running quietly on OpenRouter under the codename "Owl Alpha" before its identity was rev…updated 2026-07-212 in-linksMAI-Code-1 / MAI-Code-1-FlashMAI-Code-1-Flash: 2026-06-02 (immediately available)updated 2026-07-216 in-linksMAI-Thinking-12026-06-02 (announced at Build 2026). Exact API/GA date not announced.updated 2026-07-215 in-linksMiniMax M3June 1, 2026 — MiniMax official blog and HuggingFace release. Captured by this wiki July 17 due to WAIC 2026 …updated 2026-07-213 in-linksMistral Large 3Early access opened July 6, 2026. CEO Arthur Mensch confirmed the model on July 4, 2026 in a TechCrunch profi…updated 2026-07-212 in-linksMistral Medium 3.5Pending further detail from the full announcement materials.updated 2026-07-214 in-linksMuse ImageArena text-to-image: #2 (human-preference Elo at launch)updated 2026-07-202 in-linksMuse Spark (1.0 / 1.1)Major upgrade released alongside the Meta Model API — Meta's first paid external AI product.updated 2026-07-216 in-linksMuse VideoPreviewed July 7, 2026. General availability date not announced.updated 2026-07-201 in-linksNano Banana 2 Lite (Gemini 3.1 Flash Lite Image)June 30, 2026 — the fastest and most cost-efficient model in the Nano Banana (Gemini image generation) family…updated 2026-07-212 in-linksProject PolarisPre-announcement: 2026-06-01 (Microsoft Build 2026, June 2-3 keynote)updated 2026-07-215 in-linksQwen 3.8 Max (Preview)Preview announced July 19, 2026. Full open-weight release: no firm date.updated 2026-07-212 in-linksRobostral Navigate2026-07-08. Announced via Mistral blog (source) and covered by Bloomberg.updated 2026-07-213 in-links

Concepts

/15
Agentic Reinforcement LearningA paradigm in which an LLM agent learns via RL while interacting with an environment. Instead of a single res…23 in-linksAgents (LLM Agents)Systems that place an LLM at their core as the controller to perform multi-step planning + tool use + environ…updated 2026-06-0545 in-linksAI AlignmentThe problem of ensuring that AI systems reliably pursue goals that are beneficial to humans, and not just pro…updated 2026-07-2122 in-linksAI Control RoadmapA framework for securing AI systems at the system and infrastructure level — going beyond model-level alignme…updated 2026-07-267 in-linksAI GovernanceGovernance frameworks — legal, voluntary, and technical — that determine how frontier AI models are developed…updated 2026-07-2610 in-linksAI-Enabled CyberattacksThe use of AI models — either as intelligent assistants or fully autonomous agents — to conduct offensive cyb…updated 2026-07-2315 in-linksClaude Managed AgentsA cloud-hosted agent execution layer that separates agent logic (what Claude decides) from agent runtime (orc…updated 2026-05-235 in-linksEmbodied AgentsAgents that act through perception and actuation in the physical world or in simulated environments. Unlike p…13 in-linksGoogle ADK (Agent Development Kit)Google's Agent Development Kit (ADK) is an open-source, code-first toolkit for building, evaluating, and depl…updated 2026-07-014 in-linksGRAM — Gradient-Routed Auxiliary ModulesGRAM (Gradient-Routed Auxiliary Modules) is a modular pretraining architecture developed by Anthropic that is…updated 2026-07-212 in-linksLLM Knowledge Bases (LLM-curated personal wikis)A pattern for maintaining a continuously accumulating, structured personal or team knowledge base using an LL…updated 2026-05-165 in-linksMechanistic InterpretabilityMechanistic interpretability is the research program of reverse-engineering what specific internal computatio…updated 2026-07-204 in-linksReasoning ModelsA family of LLMs that explicitly model the reasoning process itself. They allocate test-time compute to reaso…updated 2026-05-2428 in-linksSoftware 3.0A taxonomy of software development paradigms presented by Andrej Karpathy at Sequoia Ascent 2026. A new era i…updated 2026-05-168 in-linksTest-Time Compute (Inference-Time Compute Scaling)Any technique that improves output quality by allocating additional compute at inference time. The trained mo…updated 2026-05-168 in-links

People

/7

Papers

/19
2028: Two Scenarios for Global AI Leadership — AnthropicAnthropic's policy essay argues that the US-China frontier AI gap will be decided by 2028, primarily through …updated 2026-05-182 in-linksAchieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified ScalingTBD — to be expanded in a follow-up ingest from the abstract. Title implication: olympiad-level reasoning ach…updated 2026-05-164 in-linksAgent Data Injection Attacks are Realistic Threats to AI AgentsA new attack class — Agent Data Injection (ADI) — exploits agents' trust in metadata (resource identifiers, t…updated 2026-07-141 in-linksAgentic Misalignment in Summer 2026Follow-up to the 2025 blackmail experiment series. Catalogs four new agentic misalignment failure modes acros…updated 2026-07-212 in-linksAn OpenAI model has disproved a central conjecture in discrete geometryAn OpenAI general-purpose reasoning model disproved the Erdős unit distance conjecture (planar unit distance …updated 2026-05-242 in-linksAREX: Towards a Recursively Self-Improving Agent for Deep ResearchAn agent framework that recursively improves its own research pipelines using self-evaluated quality signals,…updated 2026-07-26ASPIRE: Agentic Skills Discovery for RoboticsA continual learning system (NVIDIA GEAR Lab + UMich/UIUC/Berkeley/CMU) that lets robots autonomously write a…updated 2026-07-071 in-linksAutomated Weak-to-Strong Researcher (AAR)On an alignment research problem (weak-to-strong supervision), Anthropic's 9 AI agents achieved 97% PGR in th…updated 2026-05-315 in-linksENPIRE: Agentic Robot Policy Self-Improvement in the Real WorldA fleet of 8 real robots autonomously runs its own research loop — reading papers, proposing hypotheses, runn…updated 2026-06-275 in-linksLong-Horizon-Terminal-Bench (LHTB)A 46-task containerized terminal benchmark with dense reward grading; current best model achieves only 15.2% …updated 2026-07-151 in-linksOpenAI Parameter Golf — What It Taught UsOpenAI ran a community ML challenge (16 MB model, 10 min training, 8×H100s). Key finding: AI coding agents ha…updated 2026-05-181 in-linksPositive Alignment: Artificial Intelligence for Human FlourishingA 16-author collaborative paper from Oxford, DeepMind, Anthropic, and others. It argues that today's alignmen…updated 2026-05-313 in-linksRing-Zero: Scaling Zero RL to a Trillion Parameters for Emergent ReasoningFirst demonstration of RLVR (RL with Verifiable Rewards) scaling to 1 trillion parameters — achieves 84.2% on…updated 2026-07-191 in-linksScaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B AgentA 35B MoE model (Agents-A1) matches 1-trillion-parameter models on agentic benchmarks by scaling the agent ho…updated 2026-07-014 in-linksSEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement LearningSEED converts an LLM agent's own completed trajectories into natural-language "hindsight skills," then distil…updated 2026-07-192 in-linksSelf-Distilled Agentic Reinforcement LearningDetailed abstract to be filled in by a follow-up ingest. Inferred from the title: a self-distillation-based R…updated 2026-05-165 in-linksSolipsistic Superintelligence is Unlikely to be CooperativeCurrent RL/RLHF training treats the world as a stationary, exogenous environment. Deployed systems break this…updated 2026-06-066 in-linksThe Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement LearningLLM RL implementations use separate inference and training engines for efficiency, creating a systematic trai…updated 2026-07-173 in-linksWeak-to-Strong Generalization via Direct On-Policy DistillationRun RL on a cheap small model; transfer only the RL-induced policy delta (not the full policy) to a large mod…updated 2026-07-164 in-links