AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,504 results
9 Jul 2026

Got this setup in Claude Code to trace models routed through @merge_api! Being able to trace what your agents are doing is so valuable, I lo…

Model ReleasesDGX agent

Got this setup in Claude Code to trace models routed through @merge_api! Being able to trace what your agents are doing is so valuable, I love open source 🙌 We built a plugin that traces every Claude

GPT-5.6 is now the preferred model in Microsoft 365 Copilot

Model ReleasesDGX agent

I don't have verified information about a GPT-5.6 model or this specific announcement. This appears to be a fictional or speculative URL, as GPT-5.6 has not been publicly released by OpenAI as of my l

Here is my other heavily used pattern. Evaluator/Judge: Fable 5 Executor: GPT-5.5 I no longer wait for frontier models or am loyal to any. I…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

Here is my other heavily used pattern. Evaluator/Judge: Fable 5 Executor: GPT-5.5 I no longer wait for frontier models or am loyal to any. I now spend more time on better orchestration, harness, skill

I like choices... but now I have: 2x modes (Codex vs. Work mode) 3x GPT-5.6 models (Sol, Terra, Luna) 5x effort levels (Light, Medium, High,…

Model ReleasesDGX agent

I like choices... but now I have: 2x modes (Codex vs. Work mode) 3x GPT-5.6 models (Sol, Terra, Luna) 5x effort levels (Light, Medium, High, Extra High, Ultra) That's 2 x 3 x 5 = 30 possible configura

I'm writing a newsletter on my favorite AI models and AI tools for every single use case. This guide is specifically written for business us…

Model ReleasesDGX agent

I'm writing a newsletter on my favorite AI models and AI tools for every single use case. This guide is specifically written for business users and not engineers and will help you get the most out of

LoCA: Spatially-Aware Low-Rank Convolutional Adaptation of Vision Foundation Models

Model ReleasesDGX agent

arXiv:2607.06918v1 Announce Type: cross Abstract: Pre-trained Vision Foundation Models (VFMs) provide strong visual representations for diverse downstream tasks. The key challenge of VFM adaptation st

new meta model

Model ReleasesDGX agent

new meta model (2) Muse Spark 1.1 is strongest at agentic performance, tool use, and computer use. It does well on long-running tasks with 1M token context window, can delegate execution to sub-agents

Ollama, which helps developers run open-weight AI models locally, raised a 65M Series B led by Theory Venture, following a 15M Series A led by Benchmark (Julie Bort/TechCrunch)

Model ReleasesDGX agent

Julie Bort / TechCrunch: Ollama, which helps developers run open-weight AI models locally, raised a 65M Series B led by Theory Venture, following a 15M Series A led by Benchmark — The popular open sou

Physical activities enable scalable foundation modelling for broad-spectrum health prediction

ApplicationsDGX agent

arXiv:2607.06954v1 Announce Type: new Abstract: Wearable and mobile sensing technologies have demonstrated strong potential for health inference; however, most sensor models are designed for specific

Recursive Language Models Meet Uncertainty: The Surprising Effectiveness of Self-Reflective Program Search for Long Context

AgentsDGX agent

Long-context handling remains a core challenge for language models: even with extended context windows, models often fail to reliably extract, reason over, and use the information across long contexts

ReMoDEx: A Local-to-Global Relevance-Based Model Decision Explainability Framework for large-Scale Image Datasets

Local AiDGX agent

arXiv:2607.06889v1 Announce Type: cross Abstract: Deep learning image classifiers achieve strong predictive performance yet remain opaque in how decisions are formed. A model may predict correctly whi

Reward Valuation in Vision Language Models: Causal Mechanisms Underlying Anhedonia

Local AiDGX agent

arXiv:2607.06626v1 Announce Type: new Abstract: Recent Vision-Language Models capture increasingly complex aspects of human cognition. Here we ask whether this alignment extends to reward valuation, w

Riemannian Geometry for Pre-trained Language Model Embeddings

Model ReleasesDGX agent

arXiv:2607.07047v1 Announce Type: cross Abstract: Understanding the geometric structure of pre-trained language model embeddings matters for interpretability and safety. We ask whether sentence-level

Siamese Neural Network for Label-Efficient Critical Phenomena Prediction in 3D Percolation Models

Model ReleasesDGX agent

arXiv:2507.14159v2 Announce Type: replace-cross Abstract: Predicting critical phenomena from limited labeled data remains a challenging task in statistical physics. As percolation theory provides a ca

TF-Engram: A Train-Free Engram with SSD-Backed Memory for Large Language Models

Model ReleasesDGX agent

arXiv:2607.07388v1 Announce Type: new Abstract: Large Language Models (LLMs) store factual knowledge and domain-specific patterns implicitly in dense Transformer parameters, making knowledge expansion

Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review

ResearchDGX agent

arXiv:2607.06706v1 Announce Type: cross Abstract: Vision Language Action (VLA) models unify visual perception, natural-language understanding, and action generation within a single foundation model, a

Yesterday we launched SWE-1.7 built on the open-source Kimi K2.7. Concerns about Chinese base models are real: K2.7 completed 87% of tasks t…

AgentsDGX agent

Yesterday we launched SWE-1.7 built on the open-source Kimi K2.7. Concerns about Chinese base models are real: K2.7 completed 87% of tasks that other models refuse over human-rights concerns. We train

8 Jul 2026

Beyond the Leaderboard: A Synthesis of Tool-Use, Planning, and Reasoning Failures in Large Language Model Agents

Model ReleasesDGX agent

arXiv:2607.05775v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly evaluated on their ability to use tools, plan multi-step tasks, coordinate with other agents, and ope

Come help us scale @harvey’s model training team. If you’re interested in bringing frontier agent research into the Harvey product and worki…

AgentsDGX agent

Come help us scale @harvey’s model training team. If you’re interested in bringing frontier agent research into the Harvey product and working with: - @baseten to scale up RL to 80M+ token virtual dat

ExplAIner: A Declarative Query Language for Explaining Classification Models

ResearchDGX agent

arXiv:2607.06407v1 Announce Type: new Abstract: The XAI community has studied a wide range of queries and scores for explaining predictions of ML models. From a data management perspective, this proli

For questions that require web search, deeper reasoning, or more complex work, GPT-Live can delegate to our latest frontier model behind the…

Model ReleasesDGX agent

For questions that require web search, deeper reasoning, or more complex work, GPT-Live can delegate to our latest frontier model behind the scenes, and brings the result back into the conversation wh

I was an early tester of GPT-5.6 Sol. I was asked to not share demos until after launch but it is a very good model. It is of similar abilit…

Model ReleasesDGX agent

I was an early tester of GPT-5.6 Sol. I was asked to not share demos until after launch but it is a very good model. It is of similar ability, but quite different feel, than Fable. Fable wants to go o

Introducing GPT‑Live

Model ReleasesDGX agent

Introducing GPT‑Live OpenAI finally upgraded the model used by ChatGPT voice mode! I've had preview access for a few weeks in the iPhone app, and the new model is very impressive. It also has the abil

Learning 4D Geometric Priors for Inference-Efficient World Action Models

ApplicationsDGX agent

arXiv:2607.05468v1 Announce Type: cross Abstract: World Action Models (WAMs) have shown strong potential for robotic manipulation by jointly modeling visual future dynamics and executable action seque

Modeling Normal Is All You Need: Joint Latent Clustering for Anomaly Detection in Multimodal Cyber-Physical Systems

ResearchDGX agent

arXiv:2607.06094v1 Announce Type: new Abstract: Faults on a cyber-physical system (CPS) are too rare and unrepresentative to characterise, or even to select a model on, so detection must instead model

Parameter-Free Encoders Remain Viable for RDB Foundation Models

Model ReleasesDGX agent

arXiv:2607.05476v1 Announce Type: new Abstract: Given a relational database (RDB) storing heterogeneous tabular information, how can we predict missing (or future) values in some target column of inte

Revisiting the Relation Between Language Model Perplexity and ASR Word Error Rate for Modern End-to-End Speech Recognition

ResearchDGX agent

arXiv:2607.05612v1 Announce Type: new Abstract: Language model (LM) perplexity (PPL) has historically been used as a proxy for automatic speech recognition (ASR) word error rate (WER), with prior work

RPAM: A Principled Metric for Evaluating Associations in Language Models with High Predictive Validity in Downstream Outputs

Model ReleasesDGX agent

arXiv:2607.05679v1 Announce Type: cross Abstract: Language models (LMs) exhibit problematic biases, such as stereotypes. Effectively analyzing and mitigating such biases requires accurate and generali

These new models are rolling out to everyone over the next few days, starting today. They’ll be available in ChatGPT across iOS, Android, an…

Model ReleasesDGX agent

These new models are rolling out to everyone over the next few days, starting today. They’ll be available in ChatGPT across iOS, Android, and web. Coming soon to the API. Just tap the Voice button to

To audit SWE-Bench Pro, we used model-based investigator agents alongside independent reviews from five independent experienced software eng…

Model ReleasesDGX agent

To audit SWE-Bench Pro, we used model-based investigator agents alongside independent reviews from five independent experienced software engineers. That helped us examine tasks at scale while keeping

7 Jul 2026

20 questions for the Agentic Enterprise (and how Agent Platform can help)

Model ReleasesDGX agent

If you’re an IT leader, you might be getting a lot of questions about how to build and deploy agents. The pressure to move fast is intense, but the engineering reality is incredibly complex. Where do

CrossHallu: Do Hallucination Signals Generalize Across Languages and Domains in Large Language Model's Internals?

SafetyDGX agent

arXiv:2607.04029v1 Announce Type: new Abstract: Recent hallucination detection techniques in large language models (LLMs) focus on directly extracting features from a model's internal representations

Deep Learning Network-Temporal Models For Traffic Prediction

ResearchDGX agent

arXiv:2603.11475v2 Announce Type: replace Abstract: Accurate prediction of multivariate time series is essential for emerging network intelligent control, observability, and management functions. Exis

Diffusion Models are Open-World Affordance Learners: Leveraging Generative Priors for 3D Affordance Learning

Model ReleasesDGX agent

arXiv:2508.01651v2 Announce Type: replace Abstract: 3D affordance grounding aims to understand how diverse objects can be manipulated, making it a cornerstone of embodied interaction. However, prior w

DSWAM: A Dual-System World Action Foundation Model for Fine-Grained Robot Manipulation

SafetyDGX agent

arXiv:2607.04927v1 Announce Type: cross Abstract: World Action Models (WAMs) provide a promising alternative to Vision-Language-Action (VLA) policies by using video-based world modeling as dense super

Flex-Forcing: Towards a Unified Autoregressive and Bidirectional Video Diffusion Model

SafetyDGX agent

arXiv:2607.03509v1 Announce Type: new Abstract: Recent progress in large-scale generative models has substantially advanced video generation, yet existing methods remain constrained by a rigid inferen

GaussianArt: Unified Modeling of Geometry and Motion for Articulated Objects

Model ReleasesDGX agent

arXiv:2508.14891v3 Announce Type: replace Abstract: Reconstructing articulated objects is essential for building digital twins of interactive environments. However, prior methods typically decouple ge

GLOW-FDG: Generalized cancer LesiOn Whole-body segmentation model for ^{18}F-FDG-PET/CT

Model ReleasesDGX agent

arXiv:2607.03931v1 Announce Type: cross Abstract: Whole-body fluorodeoxyglucose positron emission tomography combined with computed tomography is widely used in cancer care, but manual lesion delineat

Harness-Aware Self-Evolving: Co-Evolving Model Weights, Harness, and Task Solutions

Model ReleasesDGX agent

arXiv:2607.03935v1 Announce Type: new Abstract: Self-evolving frameworks usually optimize task solutions while treating the surrounding harness as fixed. We introduce Harness-Aware Self-Evolving (HASE

Heaviside Continuity of Rolling Coefficients for Eliminating Epistemic Entropy in Large Language Models

AgentsDGX agent

arXiv:2607.04562v1 Announce Type: new Abstract: Large language models (LLMs) generate fluent outputs that can be wrong. Unlike humans, who often exhibit cues when providing false information, LLMs pro

Is the Geometry Doing the Work? An Operating-Point Audit of Hierarchy in Hyperbolic Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.05268v1 Announce Type: new Abstract: Whether a hyperbolic representation model uses its geometry cannot be read off its curvature parameter: what matters is the dimensionless operating poin

Language Generation with Replay: A Learning-Theoretic View of Model Collapse

ResearchDGX agent

arXiv:2603.11784v2 Announce Type: replace Abstract: As scaling laws push the training of frontier large language models (LLMs) toward ever-growing data requirements, training pipelines are approaching

MambaLIE: Scene Light Intensity-Boosted Low-Light Image Enhancement with State Space Model

Local AiDGX agent

arXiv:2607.03013v1 Announce Type: cross Abstract: Images captured by consumer electronic devices, such as mobile phones and digital cameras, often suffer from low-light degradation due to sensor limit

MCLMR: A Model-Agnostic Causal Learning Framework for Multi-Behavior Recommendation

SafetyDGX agent

arXiv:2603.25126v2 Announce Type: replace-cross Abstract: Multi-Behavior Recommendation (MBR) leverages multiple user interaction types (e.g., views, clicks, purchases) to enrich preference modeling a

Mitigating Covariate Shift in Imitation Learning for Autonomous Vehicles Using Latent Space Generative World Models

SafetyDGX agent

arXiv:2409.16663v5 Announce Type: replace-cross Abstract: We propose the use of latent space generative world models to address the covariate shift problem in autonomous driving. A world model is a ne

Model Predictive Path Integral PID Control for Learning-Based Path Following

ResearchDGX agent

arXiv:2603.29499v2 Announce Type: replace-cross Abstract: Classical proportional--integral--derivative (PID) control remains widely used in industrial control systems, while model predictive control (

Out-of-Distribution Generalization of Risk Aversion in Language Models

Model ReleasesDGX agent

arXiv:2607.02755v1 Announce Type: cross Abstract: Training AIs to be risk-averse in resources could offer a failsafe in the event that AIs turn out misaligned. Misaligned but risk-averse AIs would ten

Parameter Efficient Multimodal Instruction Tuning for Romanian Vision Language Models

Model ReleasesDGX agent

arXiv:2512.14926v2 Announce Type: replace-cross Abstract: Focusing on low-resource languages is an essential step toward democratizing generative AI. In this work, we contribute to reducing the multim

Perceptual Flow Matching for Few-Step Generative Modeling

ResearchDGX agent

arXiv:2607.03524v1 Announce Type: new Abstract: We propose Perceptual Flow Matching (PFM), a simple yet effective framework for few-step generation in flow-matching models. Rather than performing velo

PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference

ResearchDGX agent

arXiv:2511.04805v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) models have shown strong potential in scaling language models efficiently by activating only a small subset of expert

Reducing the Complexity of Deep Learning Models for EEG Analysis on Wearable Devices

Model ReleasesDGX agent

arXiv:2606.12742v3 Announce Type: replace Abstract: Wearable healthcare devices are the fastest-growing Internet of Things (IoT) sector. Many automated healthcare services rely on two crucial biologic

Reward Granularity in RLVR: Comparing Process and Outcome Reward Structures for Mathematical Reasoning in Small Language Models

SafetyDGX agent

arXiv:2607.02869v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising paradigm for improving mathematical reasoning in language models. Yet m

Sovereign AI belongs in all regions of the world. We’re working hard to make it accessible to everyone. Model weights are available on @hugg…

Model ReleasesDGX agent

Sovereign AI belongs in all regions of the world. We’re working hard to make it accessible to everyone. Model weights are available on @huggingface. Run it anywhere from corporate hardware to your per

Stage-wise Attention-Guided Region Sequencing for Adversarial Attacks on Large Vision-Language Models

Local AiDGX agent

arXiv:2602.04356v2 Announce Type: replace Abstract: Targeted adversarial attacks on Large Vision-Language Models (LVLMs) test whether small image perturbations can steer model responses toward attacke

TACG: Trajectory-Aware Commit Gating for Diffusion Language Model Decoding

Model ReleasesDGX agent

arXiv:2607.03236v1 Announce Type: new Abstract: Diffusion language models (DLLMs) generate text by iteratively denoising masked positions, exposing a trajectory of predictive distributions rather than

Taming I2V models for Image HOI Editing: A Cognitive Benchmark and Agentic Self-Correcting Framework

Model ReleasesDGX agent

arXiv:2606.19073v2 Announce Type: replace Abstract: Current image editing methods excel at static attributes but fail at complex Human-Object Interactions (HOI), a critical challenge unaddressed by ex

TestMate: Test-Time Domain Adaptation Aided by Lightweight Vision Foundation Model

Model ReleasesDGX agent

arXiv:2607.03810v1 Announce Type: new Abstract: Test-Time Domain Adaptation (TTDA) aims to adapt Deep Neural Networks to distribution shifts using only streaming, unlabeled test data in real time. Cur

World models are very interesting. https://x.com/connerruhl/status/2074245101379846400?s=20

ApplicationsDGX agent

World models are very interesting. https://x.com/connerruhl/status/2074245101379846400?s=20 Do world models work as a medium for storytelling? Over the weekend, I wanted to see how far I could push re

6 Jul 2026

Run MiniMax models on Amazon Bedrock

AgentsDGX agent

In this post, we walk through how to get started with MiniMax models on Amazon Bedrock, including the capabilities supported by these models, the service tiers available, how on-demand inference scale

4 Jul 2026

As America turns 250, we put together 250 open AI milestones from the US: open models, datasets, demos, papers, and tools that helped shape …

Model ReleasesDGX agent

As America turns 250, we put together 250 open AI milestones from the US: open models, datasets, demos, papers, and tools that helped shape the field. They go from attention is all you need, pytorch,

← Previous
1…102103104105106…1009
Next →