AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,490 results
6 Jul 2026

Claude Opus 4.8 and Sonnet 5 seem worse at tool calls than older models, likely due to post-training that assumes Claude Code-like harnesses as targets (Armin Ronacher/Armin Ronacher's Thoughts and Writings)

Model ReleasesDGX agent

Armin Ronacher / Armin Ronacher's Thoughts and Writings: Claude Opus 4.8 and Sonnet 5 seem worse at tool calls than older models, likely due to post-training that assumes Claude Code-like harnesses as

3 Jul 2026

GLM-5.2 is now selectable in Claude Code via Hugging Face🤗 Inference Providers + hf-claude. Open models are becoming easier to plug directl…

Model ReleasesDGX agent

GLM-5.2, an open-source model available through Hugging Face, can now be selected and used within Claude Code through Hugging Face Inference Providers and the hf-claude integration. This development d

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Interpreting Global Perturbation Robustness of Image Models using Axiomatic Spectral Importance Decomposition
ResearchDGX agent

arXiv:2408.01139v4 Announce Type: replace Abstract: Perturbation robustness evaluates the vulnerabilities of models, arising from a variety of perturbations, such as data corruptions and adversarial a

Meta to release new AI model with advanced coding capabilities ‘soon’

Model ReleasesDGX agent

Meta Platforms Inc. is gearing up to release a new version of its flagship Muse Spark artificial intelligence model. Alexandr Wang, the company’s chief AI officer, wrote on X today that the update wil

MMBench-Live: A Continuously Evolving Benchmark for Multimodal Models

Model ReleasesDGX agent

arXiv:2607.01813v1 Announce Type: cross Abstract: Evaluation benchmarks are essential for assessing vision-language models (VLMs), but most multimodal benchmarks are static, making them vulnerable to

OntoLearner: A Modular Python Library for Ontology Learning with Large Language Models

ResearchDGX agent

arXiv:2607.01977v1 Announce Type: new Abstract: Ontology learning (OL) aims to automatically construct structured knowledge models from text, yet progress remains fragmented across methods, domains, a

Prompt engineering is costing you money. Learn how to fine-tune your models. Stuffing a lot of text into every single API call slows down yo…

AgentsDGX agent

Prompt engineering is costing you money. Learn how to fine-tune your models. Stuffing a lot of text into every single API call slows down your app because the model has to process all those tokens bef

Psychological Imagination Networks Show Cross-Population Centrality and Clustering Alignment in Humans That Large Language Models Fail to Replicate

Model ReleasesDGX agent

arXiv:2510.04391v5 Announce Type: replace Abstract: Mental imagery vividness is a stable individual trait, yet whether imagined scenarios share relational structure across human and synthetic large la

Sources: Alibaba has banned employees from using Claude Code and asked them to remove all Claude models from their work computers, citing security concerns (The Information)

Model ReleasesDGX agent

The Information: Sources: Alibaba has banned employees from using Claude Code and asked them to remove all Claude models from their work computers, citing security concerns — Alibaba Group has banned

The Wiola Architecture for Efficient Small Language Models

Model ReleasesDGX agent

arXiv:2607.01394v1 Announce Type: new Abstract: We present Wiola, a fully original Small Language Model (SLM) architecture built from first principles, sharing no structural lineage with any existing

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as…

Model ReleasesDGX agent

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as a Google replacement, for homework “help,” etc. It is someo

Unpopular opinion: While everyone is so hyped about Fable, GPT5.6 and other huge and expensive models, I think the real hero of the last few…

Model ReleasesDGX agent

Unpopular opinion: While everyone is so hyped about Fable, GPT5.6 and other huge and expensive models, I think the real hero of the last few months is *Qwen 27b*. Our ML/AI engineering teams are have

2 Jul 2026

3D Point World Models: Point Completion Enables More Accurate Dynamics Learning

ResearchDGX agent

arXiv:2607.00148v1 Announce Type: cross Abstract: Learning predictive models of the world enables robotic control through planning, potentially allowing robots to improvise solutions on new tasks. How

AD-MPCC: Adaptive Differentiable Model Predictive Contouring Control for Autonomous Racing

Model ReleasesDGX agent

arXiv:2607.00141v1 Announce Type: new Abstract: This paper presents Adaptive Differentiable Model Predictive Contouring Control (AD-MPCC), a framework for autonomous racing that integrates differentia

Device Passport: Enabling Spatio-Temporal Pretrained Models to Generalize Across Input Layouts

ResearchDGX agent

arXiv:2607.00249v1 Announce Type: new Abstract: New device layouts pose a challenging modeling problem due to the lack of large datasets for each specific layout. Biosignal foundation models offer a p

DriveVA: Video Action Models are Zero-Shot Drivers

Model ReleasesDGX agent

arXiv:2604.04198v2 Announce Type: replace Abstract: Generalization is a central challenge in autonomous driving, as real-world deployment requires robust performance under unseen scenarios, sensor dom

EmbodimentSemantic: A Spatial Scene-Graph Dataset and Benchmark for Vision-Language Models on Embodied Manipulation Trajectories

Model ReleasesDGX agent

arXiv:2607.00020v1 Announce Type: new Abstract: Spatial grounding remains a key limitation of vision-language-action (VLA) systems for robotic manipulation. While current models can recognize objects

Foundation Models vs. Radiomics for Lung Computed Tomography: A Benchmark of Feature Extractors, Classification Heads, and Segmentation Choices

Model ReleasesDGX agent

arXiv:2607.01001v1 Announce Type: new Abstract: Radiomics is the established approach for CT-based lung cancer phenotyping, yet comparisons with foundation models rarely isolate contributions of featu

Generative Modeling of Quantum Distribution with Functional Flow Matching

TutorialsDGX agent

arXiv:2607.00301v1 Announce Type: new Abstract: The emergence of powerful deep generative models based on diffusion and flow matching has enabled the learning and modeling of complex distributions. Le

Hey, That's My Model! Introducing Chain & Hash, An LLM Fingerprinting Technique

ResearchDGX agent

arXiv:2407.10887v4 Announce Type: replace-cross Abstract: Growing concerns over the theft and misuse of Large Language Models (LLMs) underscore the need for effective fingerprinting to link a model to

Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models

ResearchDGX agent

arXiv:2607.01222v1 Announce Type: new Abstract: Recent 3D generative models can synthesize high-quality geometry but often struggle to reproduce intricate textures from reference images, largely due t

MetaOthello: A Controlled Study of Multiple World Models in Transformers

TutorialsDGX agent

arXiv:2602.23164v2 Announce Type: replace Abstract: Foundation models must handle multiple generative processes, yet mechanistic interpretability largely studies capabilities in isolation; it remains

Prototype Language Models

Local AiDGX agent

arXiv:2607.00510v1 Announce Type: new Abstract: Knowing which training examples drive outputs is fundamental to auditing, correcting, and understanding language models, yet for modern LLMs this remain

Sources: Alexandr Wang said Meta's model currently in training, codenamed Watermelon, matches GPT-5.5 and uses an 'order of magnitude more compute than Avocado' (Business Insider)

Model ReleasesDGX agent

Business Insider: Sources: Alexandr Wang said Meta's model currently in training, codenamed Watermelon, matches GPT-5.5 and uses an “order of magnitude more compute than Avocado” — Meta is making sign

Unleashing More Actions via Action Compositional Training for VLA Models

SafetyDGX agent

arXiv:2607.00351v1 Announce Type: new Abstract: Vision-Language-Action models excel at robotic manipulation, driven by the scale and diversity of demonstration data. However, standard training paradig

When Less is More: 8-bit Quantization Improves Continual Learning in Large Language Models

ResearchDGX agent

arXiv:2512.18934v2 Announce Type: replace-cross Abstract: Catastrophic forgetting poses a fundamental challenge in continual learning, particularly when models are quantized for deployment efficiency.

1 Jul 2026

A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks, Risks, Defenses, and Open Problems

SafetyDGX agent

arXiv:2606.31639v1 Announce Type: cross Abstract: Large language models are no longer only text generators. They are increasingly embedded in retrieval pipelines, enterprise assistants, coding environ

Anthropic launches Claude Sonnet 5 AI model with coding, safety upgrades as Fable and Mythos controls lifted

Model ReleasesDGX agent

Anthropic PBC today debuted Claude Sonnet 5, a midrange large language model that outperforms its predecessor in several areas. The LLM will be the default option in the consumer tiers of the company’

Ask the World Before Acting: Budgeted Environment Probing for World-Model Calibration

SafetyDGX agent

arXiv:2606.31422v1 Announce Type: new Abstract: Long-horizon language agents do not only choose actions; they carry a private model of the world from one decision to the next. When that model drifts,

Benchmarking Large Language Models on Floating-Point Error Classification

Model ReleasesDGX agent

arXiv:2606.31308v1 Announce Type: new Abstract: This paper investigates the capability of Large Language Models (LLMs) to detect and classify floating-point errors statically in software code. We intr

Deductive Logic in Language Models: Horizontal vs Vertical Reasoning

TutorialsDGX agent

arXiv:2510.09340v2 Announce Type: replace Abstract: Recent language models exhibit significant logical reasoning abilities, yet the mechanisms supporting deductive inference remain poorly understood.

Deep Spectral Models for Robust Dental Shape Generation

ApplicationsDGX agent

arXiv:2606.31293v1 Announce Type: new Abstract: Accurate modeling of dental crown morphology is fundamental for diagnosis, orthodontic planning, and computer-aided restoration design. However, dataset

DVG-WM: Disentangled Video Generation Enables Efficient Embodied World Model for Robotic Manipulation

ApplicationsDGX agent

arXiv:2606.32028v1 Announce Type: new Abstract: Video-based embodied world models provide an appealing substrate for robotic manipulation by predicting future states, yet current approaches remain lim

E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes

Model ReleasesDGX agent

arXiv:2604.04834v2 Announce Type: replace Abstract: Robotic Vision-Language-Action (VLA) models generalize well for open-ended manipulation, but their perception is fragile under sensing-stage degrada

Neurologyca launches research labs to help frontier AI models understand people

Model ReleasesDGX agent

Neurologyca, a company working on human-centric artificial intelligence, announced the launch of a research division on Tuesday dedicated to advancing AI models’ ability to understand and adapt to peo

Optimization Algorithms for Joint OFDM Waveform Design and RIS Configuration in 6G Networks: From Convex Relaxation to Foundation Models

Model ReleasesDGX agent

arXiv:2606.31334v1 Announce Type: new Abstract: Joint OFDM-RIS optimization for 6G is a mixed-integer nonlinear programming (MINLP) problem covering sum-rate maximization, energy efficiency, max-min f

Position: Vision-Language-Action Models Cannot Be Verified to Perform Physical Reasoning

Model ReleasesDGX agent

arXiv:2606.30686v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) systems, built on pretrained vision-language models (VLMs), have shown rapidly improving performance on robot manipulatio

The Harness should be Model agnostic. The Knowledge it builds around you is your moat. You should not have to migrate or give that up just b…

AgentsDGX agent

The Harness should be Model agnostic. The Knowledge it builds around you is your moat. You should not have to migrate or give that up just because a model is banned or just not good. Right now in my R

Towards Inclusive Mobility Modeling: Characterizing and Evaluating Elderly Trajectory Patterns in Urban Systems

Local AiDGX agent

arXiv:2606.31207v1 Announce Type: new Abstract: The rapid advance of smart cities increasingly depends on trajectory data mining, yet underrepresented demographic groups, particularly the elderly, are

Unified Structural-Hydrodynamic Modeling of Underwater Underactuated Mechanisms and Soft Robots

Model ReleasesDGX agent

arXiv:2603.07939v2 Announce Type: replace Abstract: Underwater robots are widely deployed for ocean exploration and manipulation. Underactuated mechanisms are advantageous in aquatic environments beca

30 Jun 2026

AERIS: Aerial-Edge Role-Driven Intelligence at Runtime via Orchestrated Language-Model Swarm

Model ReleasesDGX agent

arXiv:2606.30151v1 Announce Type: new Abstract: Integrating large language models into robotic systems holds promise for enhancing autonomy, yet practical deployment remains constrained by strict hear

Bridgewater, one of the worlds largest hedge funds, a Tinker customer talks through how they've carefully fine-tuned a model focused on what…

SafetyDGX agent

Bridgewater, one of the worlds largest hedge funds, a Tinker customer talks through how they've carefully fine-tuned a model focused on what makes interesting financial news. Their fine-tuned model is

CASE-Bench: Context-Aware SafEty Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2501.14940v4 Announce Type: replace-cross Abstract: Aligning large language models (LLMs) with human values is essential for their safe deployment and widespread adoption. Current LLM safety ben

DiLaServe: High SLO Attainment Serving for Diffusion Language Models

ApplicationsDGX agent

arXiv:2606.29094v1 Announce Type: new Abstract: Diffusion language models (DLMs) have recently emerged as a promising alternative to conventional autoregressive language models. By generating multiple

Google launches Nano Banana 2 Lite, a low cost text-to-image model that delivers outputs in four seconds, and rolls out Gemini Omni Flash to developers (The Keyword)

Model ReleasesDGX agent

The Keyword: Google launches Nano Banana 2 Lite, a low cost text-to-image model that delivers outputs in four seconds, and rolls out Gemini Omni Flash to developers — We're making it easier to experim

How to Train Your Long-Context Visual Document Model

Model ReleasesDGX agent

arXiv:2602.15257v3 Announce Type: replace-cross Abstract: We present the first comprehensive, large-scale study of training long-context vision language models up to 344K context, targeting long-docum

Insidious by Design: Implications of Large Language Model algorithmic bias for the Global South

Model ReleasesDGX agent

arXiv:2606.28333v1 Announce Type: cross Abstract: egin{quote} The biases in Large Language Models' (LLMs) outputs remain inadequately theorised, particularly from the perspective of the Global South.

J-LAW: Joint Localization and Actionable World Modeling via Coupled Latent Factor Graphs

Local AiDGX agent

arXiv:2606.28712v1 Announce Type: cross Abstract: Classical SLAM estimates metric poses and a geometric map but produces no actionable predictive model for planning. Action-conditioned world models le

Legal Domain Adaptation of Modern BERT Models

ApplicationsDGX agent

arXiv:2606.28538v1 Announce Type: new Abstract: We investigate domain adaptation of modern BERT models in the legal domain. We further pre-train ModernBERT on all US court opinions using the masked la

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models

Model ReleasesDGX agent

arXiv:2601.05366v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed as agents that invoke external tools through structured function calls. While recent wo

Love how Google continues to drive down the cost of building with their models. <4s image and $0.034 / 1K image. Wow! We have a bunch of stu…

Model ReleasesDGX agent

Love how Google continues to drive down the cost of building with their models. <4s image and 0.034 / 1K image. Wow! We have a bunch of stuff (education & research) we're building @dair_ai using Nano

One Scene, Two Depths: Probing Geometric Ambiguity in Monocular Foundation Models

Model ReleasesDGX agent

arXiv:2606.29600v1 Announce Type: cross Abstract: A faithful 3D world representation should account for layered geometry, where a single camera ray may contain multiple visible and geometrically valid

Online Experiential Learning for Language Models

SafetyDGX agent

arXiv:2603.16856v2 Announce Type: replace Abstract: The prevailing paradigm for improving large language models relies on offline training with human annotations or simulated environments, leaving the

Phonological Perception of Sign Language Models

SafetyDGX agent

arXiv:2606.28667v1 Announce Type: new Abstract: Sign languages are compositional systems where meaning arises by combining sublexical phonological parameters, such as handshape, location, and movement

Physics Models for Sim-to-Real Transfer in Professional-Level Robot Table Tennis

SafetyDGX agent

arXiv:2606.28805v1 Announce Type: new Abstract: At competitive speeds and spins, a table tennis ball follows complex, counterintuitive trajectories that a robot must track and precisely counter within

Reliability, Faithfulness, and the Limits of Post-hoc Explanations of Opaque Scientific Models

ResearchDGX agent

arXiv:2606.29346v1 Announce Type: new Abstract: Post-hoc explanation methods are routinely used to interpret scientific machine learning models, with the deliverable understood to be insight into the

SAKE: Software Architectural Knowledge Evaluation Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2606.29520v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used as assistants across the software development lifecycle, yet their ability to reason about software

ScAle: Attention Head Scaling as a Minimal Adapter for Spatial Reasoning in Vision Language Models

Model ReleasesDGX agent

arXiv:2606.29579v1 Announce Type: cross Abstract: Spatial reasoning remains a persistent challenge for many vision language models (VLMs), and improving it typically requires fine-tuning with substant

Sonnet 5 is here! This is going to support better long-running agents. Previous Sonnet models were unreliable, so it's great to see the impr…

Model ReleasesDGX agent

Sonnet 5 is here! This is going to support better long-running agents. Previous Sonnet models were unreliable, so it's great to see the improved version that can complete agentic tasks more reliably.

Super excited about open-source router systems and routing models like @vllm_project semantic router: https://huggingface.co/llm-semantic-ro…

IndustryDGX agent

Super excited about open-source router systems and routing models like @vllm_project semantic router: https://huggingface.co/llm-semantic-router The future is multi-models and you'll want to customize

← Previous
1…8182838485…1009
Next →