AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,563 results
9 Jul 2026

Geometric Self-Distillation for Reasoning Generalization

Model ReleasesDGX agent

arXiv:2607.06855v1 Announce Type: cross Abstract: On-policy distillation is a practical post-training recipe for large language models, supplying dense teacher supervision on the student's own traject

GeoProp: Grounding Robot State in Vision for Generalist Manipulation

Model ReleasesDGX agent

arXiv:2607.07101v1 Announce Type: cross Abstract: Proprioception is fundamental to robotic manipulation, yet standard fusion methods often treat it as an isolated vector lacking explicit alignment wit

GIFT: Geometry-Informed Low-precision Gradient Communication for LLM Pretraining

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.07494v1 Announce Type: cross Abstract: Gradient communication is a primary scaling bottleneck in large language model (LLM) pretraining. Communicating gradients in low-precision formats, su

GLM 5.2: a new rise of open-weight agentic models

Model ReleasesDGX agent

On June 16th, Z.ai released GLM 5.2, its latest flagship model. At the time of announcement, it advertised scores at or near Anthropic and OpenAI's models, and far ahead of GLM 5.1. In the world of us

Got this setup in Claude Code to trace models routed through @merge_api! Being able to trace what your agents are doing is so valuable, I lo…

Model ReleasesDGX agent

Got this setup in Claude Code to trace models routed through @merge_api! Being able to trace what your agents are doing is so valuable, I love open source 🙌 We built a plugin that traces every Claude

GPT-5.6: Frontier intelligence that scales with your ambition

Model ReleasesDGX agent

GPT-5.6 is OpenAI's advanced language model featuring improved scalability and performance capabilities. The model is designed to handle increasingly complex tasks and larger-scale applications, adapt

GPT-5.6 is now available in Devin! On FontierCode 1.1 Extended, the GPT-5.6 family stands out for pairing strong scores with excellent cost …

Model ReleasesDGX agent

GPT-5.6 is now available in Devin! On FontierCode 1.1 Extended, the GPT-5.6 family stands out for pairing strong scores with excellent cost efficiency. GPT 5.6 Sol reaches top performance at nearly ha

GPT-5.6 is now supported in Hermes Agent and available via Nous Portal

Model ReleasesDGX agent

Nous Research announced support for GPT-5.6 integration within Hermes Agent, with access available through the Nous Portal platform. This update enables users to leverage the capabilities of GPT-5.6 t

GPT-5.6 is now the preferred model in Microsoft 365 Copilot

Model ReleasesDGX agent

I don't have verified information about a GPT-5.6 model or this specific announcement. This appears to be a fictional or speculative URL, as GPT-5.6 has not been publicly released by OpenAI as of my l

GPT-5.6 Sol costs 5 per 1M input tokens and 30 per 1M output tokens, GPT-5.6 Terra costs 2.50 and 15, and GPT-5.6 Luna costs 1 and 6 (OpenAI)

Model ReleasesDGX agent

OpenAI: GPT-5.6 Sol costs 5 per 1M input tokens and 30 per 1M output tokens, GPT-5.6 Terra costs 2.50 and 15, and GPT-5.6 Luna costs 1 and 6 — More intelligence from every token, stronger performance

GPT-5.6 Sol sets a new SOTA on ARC-AGI-3: 7.8% Sol is the first verified frontier model to ever beat an ARC-AGI-3 game It is the best model …

Model ReleasesDGX agent

GPT-5.6 Sol achieved a breakthrough by becoming the first verified frontier AI model to surpass performance on ARC-AGI-3, scoring 7.8% and setting a new state-of-the-art benchmark. ARC-AGI (Abstractio

GPT-5.6 Sol, Terra, and Luna are now available in Cursor. On CursorBench, Sol scores 67.2%.

Model ReleasesDGX agent

Cursor has released three new AI models named Sol, Terra, and Luna, with Sol achieving a 67.2% score on CursorBench. These models are now available for use within the Cursor development environment. T

GPT-Live is now fully rolled out to all ChatGPT users on Go, Plus, and Pro plans. Free user rollout is in progress. Update to the latest ver…

Model ReleasesDGX agent

GPT-Live is now fully rolled out to all ChatGPT users on Go, Plus, and Pro plans. Free user rollout is in progress. Update to the latest version of the ChatGPT app on iOS or Android to try it out. Int

GPT‑5.6 is available starting today across ChatGPT, Codex, and the OpenAI API. The rollout is starting globally now and will continue gradua…

Model ReleasesDGX agent

GPT‑5.6 is available starting today across ChatGPT, Codex, and the OpenAI API. The rollout is starting globally now and will continue gradually toward full availability over the next 24 hours. In Chat

GrandTour: A Legged Robotics Dataset in the Wild for Multi-Modal Perception and State Estimation

Model ReleasesDGX agent

arXiv:2602.18164v3 Announce Type: replace Abstract: Accurate state estimation and multi-modal perception are prerequisites for autonomous legged robots in complex, large-scale environments. To date, n

Grok 4.5 is also rank 1 in SWE marathon

Model ReleasesDGX agent

Grok 4.5, xAI's large language model, has achieved the top ranking in the SWE (Software Engineering) Marathon benchmark, according to an announcement by Elon Musk. This ranking suggests the model demo

Grok 4.5 is dominating the latest AI leaderboards Claims the #1 spot: • #1 on AutomationBench-AA • #1 on Terminal-Bench v2 • #1 on Harvey Le…

Model ReleasesDGX agent

Grok 4.5 is dominating the latest AI leaderboards Claims the #1 spot: • #1 on AutomationBench-AA • #1 on Terminal-Bench v2 • #1 on Harvey Legal Agent Benchmark • #1 on SWE Marathon • #1 on SWE-Atlas-Q

Grok 4.5 is now ranked #1 on τ³-Banking in Artificial Analysis Ahead of GPT-5.5 xhigh, Claude Fable 5 and Claude 4.8 (max)

Model ReleasesDGX agent

I cannot provide a factual summary for this entry as the URL, date stamp, and specific benchmark claims cannot be verified. The title references AI models and rankings that do not appear to correspond

grok 4.5 made me give grok build a serious run today here's my honest first impression (non affiliated neutral view point): 1. grok build is…

Model ReleasesDGX agent

grok 4.5 made me give grok build a serious run today here's my honest first impression (non affiliated neutral view point): 1. grok build is a very good harness firstmate stretches harness capabilitie

Grok Build mogs Claude Code & Codex TUIs i can't lie. Not to mention the perf of Grok 4.5... It's not even a competition at this point. Good…

Model ReleasesDGX agent

This post compares Grok's code generation capabilities favorably against Claude Code and OpenAI's Codex, claiming superior performance particularly in Grok 4.5's speed and functionality. The author as

Grok x Cursor

Model ReleasesDGX agent

Grok, Elon Musk's AI assistant developed by xAI, has integrated with Cursor, a popular AI-powered code editor, enabling developers to leverage Grok's capabilities for coding tasks and assistance. This

Grounding Spatial Relations in a Compact World Model: Instruction Leakage and a Goal-Free Dynamics Fix

Model ReleasesDGX agent

arXiv:2607.06925v1 Announce Type: new Abstract: Compact world models that condition on a language goal promise to ground relations such as ``put the red block left of the blue block'' using a sparse s

HAJJv2-CrowdCount: Zero-Shot Benchmark for Dense Crowd Counting

Model ReleasesDGX agent

arXiv:2607.07322v1 Announce Type: cross Abstract: Automated crowd counting in Hajj video is difficult not because current models lack capacity, but because the footage violates the assumptions those m

Harrison Chase @hwchase17: Nemotron 3 Ultra hit 86%. Claude Opus hit 87%. At one-tenth the cost. Chase runs LangChain and disclosed it on st…

Model ReleasesDGX agent

Harrison Chase @hwchase17: Nemotron 3 Ultra hit 86%. Claude Opus hit 87%. At one-tenth the cost. Chase runs LangChain and disclosed it on stage: inside LangChain's internal deep-agents benchmark, open

Health System Scale Semantic Search Across Unstructured Clinical Notes

Model ReleasesDGX agent

arXiv:2604.25605v2 Announce Type: replace-cross Abstract: Introduction: Semantic search, which retrieves documents based on conceptual similarity rather than keywords, offers advantages for retrieval

Healthier LLMs: Retrieval-Augmented Generation for Public Health Question Answering

Model ReleasesDGX agent

arXiv:2607.06641v1 Announce Type: cross Abstract: Large language models (LLMs) achieve promising results on medical question answering benchmarks, yet their use in public health is constrained by hall

Here is my other heavily used pattern. Evaluator/Judge: Fable 5 Executor: GPT-5.5 I no longer wait for frontier models or am loyal to any. I…

Model ReleasesDGX agent

Here is my other heavily used pattern. Evaluator/Judge: Fable 5 Executor: GPT-5.5 I no longer wait for frontier models or am loyal to any. I now spend more time on better orchestration, harness, skill

Higher-Order Geometric Updates for Levenberg-Marquardt Method via Riemann Normal Coordinates

Model ReleasesDGX agent

arXiv:2607.07623v1 Announce Type: new Abstract: Nonlinear least-squares optimization is central to regression, physics-informed neural networks, and other machine-learning tasks. Such problems have a

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Model ReleasesDGX agent

arXiv:2506.08797v2 Announce Type: replace Abstract: To address key limitations in human-object interaction (HOI) video generation -- specifically the reliance on curated motion data, limited generaliz

i do love rottweilers

Model ReleasesDGX agent

i do love rottweilers My view of: Fable 5 vs GPT-5.6-Sol. They are not easy models to compare, these are my vibes - take them as you will. My overall feel is that Fable is a 'wise owl' who is very tho

I like choices... but now I have: 2x modes (Codex vs. Work mode) 3x GPT-5.6 models (Sol, Terra, Luna) 5x effort levels (Light, Medium, High,…

Model ReleasesDGX agent

I like choices... but now I have: 2x modes (Codex vs. Work mode) 3x GPT-5.6 models (Sol, Terra, Luna) 5x effort levels (Light, Medium, High, Extra High, Ultra) That's 2 x 3 x 5 = 30 possible configura

I'm writing a newsletter on my favorite AI models and AI tools for every single use case. This guide is specifically written for business us…

Model ReleasesDGX agent

I'm writing a newsletter on my favorite AI models and AI tools for every single use case. This guide is specifically written for business users and not engineers and will help you get the most out of

Imputation Meets Clustering: Exploiting Latent Subgroup Structure for Missing Data Recovery

Model ReleasesDGX agent

arXiv:2607.06930v1 Announce Type: cross Abstract: Missing data is prevalent in practical applications, making effective imputation an essential preprocessing step for downstream analysis. Real-world d

Information Allocation Dynamics in Neural Network Optimization

Model ReleasesDGX agent

arXiv:2607.07156v1 Announce Type: new Abstract: Different optimizers have different update biases, but these biases are usually implicit. Existing studies mainly analyze or control such biases from th

InfraQR: Edge-Placed QR-Inspired Structured Patch Attacks on Infrared Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.07288v1 Announce Type: new Abstract: Infrared vision-language models are increasingly used for perception under low-light and adverse visual conditions, yet their robustness to localized st

Institutional Red-Teaming: Deployment Rules, Not Just Models, Causally Shape Multi-Agent AI Safety

Model ReleasesDGX agent

arXiv:2607.07695v1 Announce Type: new Abstract: We introduce institutional red-teaming, an evaluation methodology for testing deployment rules in multi-agent AI: hold the agents, objectives, and task

Interesting comparison of Grok & Opus 1M+ context window coming soon

Model ReleasesDGX agent

Elon Musk announced an upcoming comparison between Grok and Anthropic's Claude Opus with its 1M+ token context window. The post suggests xAI plans to demonstrate how Grok's capabilities compare to Cla

Is Meta AI back? Haven't seen Mark post in three years here. Plus, the model is available via API. Not to mention the courage to announce it…

Model ReleasesDGX agent

Is Meta AI back? Haven't seen Mark post in three years here. Plus, the model is available via API. Not to mention the courage to announce it the same week as the long-awaited GPT-5.6. Great timing if

It all comes together in 10 minutes. http://openai.com/live

Model ReleasesDGX agent

OpenAI announced a live event or demonstration where key information or announcements would be revealed within a 10-minute timeframe. This post appears to reference a live stream or real-time presenta

It all started when I knocked on a door in Palo Alto. I saw a little llama icon on the door. Michael opened it. That's how I became friends …

Model ReleasesDGX agent

It all started when I knocked on a door in Palo Alto. I saw a little llama icon on the door. Michael opened it. That's how I became friends with the Ollama founders. Today we led their $65m Series B.

Just shared my favorite AI model and tool for every use case. Check out the list here: https://aiwithallie.beehiiv.com/p/the-best-ai-model-a…

Model ReleasesDGX agent

Just shared my favorite AI model and tool for every use case. Check out the list here: https://aiwithallie.beehiiv.com/p/the-best-ai-model-and-tool-for-every-use-case I'm writing a newsletter on my fa

langsmith for coding agents

Model ReleasesDGX agent

langsmith for coding agents We built a plugin that traces every Claude Code session straight into LangSmith. Three commands, one JSON block, and every message, tool call, and subagent run shows up as

Large Language Models (LLMs) and Generative AI in Cybersecurity and Privacy: A Survey of Dual-Use Risks, AI-Generated Malware, Explainability, and Defensive Strategies

Model ReleasesDGX agent

arXiv:2607.06963v1 Announce Type: cross Abstract: Large Language Models (LLMs) and generative AI (GenAI) systems, such as ChatGPT, Claude, Gemini, LLaMA, Copilot, Stable Diffusion by OpenAI, Anthropic

Latency-Aware Bid Acceptance under Operational Feasibility: A Public Benchmark with Hindsight Ceilings

Model ReleasesDGX agent

arXiv:2607.07343v1 Announce Type: cross Abstract: Online truckload bid acceptance is a closed-loop stochastic decision problem in which a carrier or broker must, in real time, accept or reject a tende

LiveOIBench: Can Large Language Models Outperform Human Contestants in Informatics Olympiads?

Model ReleasesDGX agent

arXiv:2510.09595v3 Announce Type: replace Abstract: Competitive programming problems are increasingly used to evaluate the coding capabilities of large language models (LLMs) due to their complexity a

LoCA: Spatially-Aware Low-Rank Convolutional Adaptation of Vision Foundation Models

Model ReleasesDGX agent

arXiv:2607.06918v1 Announce Type: cross Abstract: Pre-trained Vision Foundation Models (VFMs) provide strong visual representations for diverse downstream tasks. The key challenge of VFM adaptation st

Looks like Grok 4.5 is #1 on at least a few benchmarks. Better than expected.

Model ReleasesDGX agent

Elon Musk announced that Grok 4.5, an AI model developed by xAI, has achieved top performance on several benchmarks, exceeding prior expectations. The post suggests the model's capabilities have surpa

MADB: A Large-Scale Music Aesthetics Dataset with Professional and Multi-Dimensional Annotations

Model ReleasesDGX agent

arXiv:2607.06929v1 Announce Type: cross Abstract: Music aesthetic assessment is a challenging yet underexplored problem, requiring models to capture fine-grained, multi-dimensional human perceptual ju

Making Implicit Preservation Intent Explicit in Conversational Image Editing

Model ReleasesDGX agent

arXiv:2607.07051v1 Announce Type: cross Abstract: Conversational image editing requires preserving not only visible content, but also content that temporarily disappears across turns. When newly added

Massive day for us @OpenAI: - GPT-5.6 SOTA at ~everything & by far most token efficient - Agents for everyone in the new ChatGPT app Work an…

Model ReleasesDGX agent

Massive day for us @OpenAI: - GPT-5.6 SOTA at ~everything & by far most token efficient - Agents for everyone in the new ChatGPT app Work and Codex modes - Work mode available on desktop (most powerfu

Measuring the metacognition of AI

Model ReleasesDGX agent

arXiv:2603.29693v3 Announce Type: replace Abstract: A robust decision-making process must take into account uncertainty, especially when the choice involves inherent risks. Because artificial intellig

MedPMC: A Systematic Framework for Scaling High-Fidelity Medical Multimodal Data for Foundation Models

Model ReleasesDGX agent

arXiv:2607.07673v1 Announce Type: new Abstract: Medicine is inherently multimodal, requiring clinicians to synthesize information across diverse data streams. Yet the development of multimodal foundat

Meet Bartosz (@nasqret). A mathematician solving previously unsolvable math problems with GPT-5.6

Model ReleasesDGX agent

Bartosz (@nasqret) is a mathematician who uses GPT-5.6 to solve mathematical problems that were previously considered unsolvable or intractable. This case study from OpenAI highlights how advanced lar

Meet Hiroki (@tomiyasu16). A broccoli farmer running his farm with GPT-5.6.

Model ReleasesDGX agent

Hiroki is a broccoli farmer who utilizes GPT-5.6 to manage and optimize his farm operations. This case study demonstrates practical applications of advanced AI language models in agricultural manageme

Meet the Wishingrads. A family running a cereal business from their dining room with GPT-5.6.

Model ReleasesDGX agent

The Wishingrads are a family operating a cereal business from their home using GPT-5.6, likely demonstrating how advanced AI language models can assist small entrepreneurs and home-based businesses wi

Meta launches a Meta Model API, which Mark Zuckerberg says will have 'aggressive and attractive' pricing at ~25% of the cost of OpenAI's and Anthropic's models (Kurt Wagner/Bloomberg)

Model ReleasesDGX agent

Kurt Wagner / Bloomberg: Meta launches a Meta Model API, which Mark Zuckerberg says will have “aggressive and attractive” pricing at ~25% of the cost of OpenAI's and Anthropic's models — In a crowded

Meta launches flagship Muse Spark 1.1 model with multi-agent upgrades

Model ReleasesDGX agent

Meta Platforms Inc. today launched a new flagship large language model optimized to power multi-agent automation workflows. Muse Spark 1.1 is available in the company’s Meta AI chatbot service and via

Meta says its new AI model is ready to compete on coding

Model ReleasesDGX agent

After reentering the AI race with its first in-house Muse Spark model in April, Meta is now opening up the doors to developers with a new model that can plug into AI coding software with the new Meta

MIRA-Math: A Benchmark for Minimal Information Requesting and Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2607.07391v1 Announce Type: new Abstract: Mathematical reasoning benchmarks typically provide all facts needed to solve each problem, while interactive benchmarks often mix reasoning with tools,

Multi-Agent Robotic Control with Onboard Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.07403v1 Announce Type: cross Abstract: Vision Language Models (VLMs) and Vision Language Action (VLA) models have shown promise in robotic control. Yet, they face significant challenges reg

← Previous
1…9091929394…377
Next →