AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,405 results
Model Releases

Model-Driven Policy Optimization in Differentiable Simulators via Stochastic Exploration

DGX agent

arXiv:2605.07520v1 Announce Type: new Abstract: Differentiable planning enables gradient-based optimization of decision-making problems by leveraging differentiable models of system dynamics. However,

model-releasesarxiv-cs-ai
11 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

OmicsLM: A Multimodal Large Language Model for Multi-Sample Omics Reasoning

DGX agent

arXiv:2605.06728v1 Announce Type: cross Abstract: Interpreting transcriptomic data is one of the most common analytical tasks in modern biology. Yet most current models either consume expression profi

model-releasesarxiv-cs-ai
11 May 2026
Safety

Self-Programmed Execution for Language-Model Agents

DGX agent

arXiv:2605.06898v1 Announce Type: new Abstract: At the heart of existing language model agents is a fixed orchestrator program responsible for the state transition between consecutive turns. This pape

safetyarxiv-cs-ai
11 May 2026
Local Ai

TAP: Two-Stage Adaptive Personalization of Multi-Task and Multi-Modal Foundation Models in Federated Learning

DGX agent

arXiv:2509.26524v3 Announce Type: replace-cross Abstract: In federated learning (FL), local personalization of models has received significant attention, yet personalized fine-tuning of foundation mod

local-aiarxiv-cs-ai
11 May 2026
Model Releases

Teaching Language Models to Think in Code

DGX agent

arXiv:2605.07237v1 Announce Type: new Abstract: Tool-integrated reasoning (TIR) has emerged as a dominant paradigm for mathematical problem solving in language models, combining natural language (NL)

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Tool Calling is Linearly Readable and Steerable in Language Models

DGX agent

arXiv:2605.07990v1 Announce Type: cross Abstract: When a tool-calling agent picks the wrong tool, the failure is invisible until execution: the email gets sent, the meeting gets missed. Probing 12 ins

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Uneven Evolution of Cognition Across Generations of Generative AI Models

DGX agent

arXiv:2605.06815v1 Announce Type: new Abstract: The pursuit of artificial general intelligence necessitates robust methods for evaluating the cognitive capabilities of models beyond narrow task perfor

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

When Are Experts Misrouted? Counterfactual Routing Analysis in Mixture-of-Experts Language Models

DGX agent

arXiv:2605.07260v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) language models route each token to a small subset of experts, but whether the routes selected by a trained top-k router are

model-releasesarxiv-cs-cl
11 May 2026
Research

InSpatio-WorldFM: An Open-Source Real-Time Generative Frame Model

DGX agent

arXiv:2603.11911v3 Announce Type: replace Abstract: We present InSpatio-WorldFM, an open-source real-time frame model for spatial intelligence. Unlike video-based world models that rely on sequential

researcharxiv-cs-cv
7 May 2026
Model Releases

Not All That Is Fluent Is Factual: Investigating Hallucinations of Large Language Models in Academic Writing

DGX agent

arXiv:2605.04171v1 Announce Type: new Abstract: Large Language models (LLMs) show extraordinary abilities, but they are still prone to hallucinations, especially when we use them for generating Academ

model-releasesarxiv-cs-cl
7 May 2026
Hardware

Piper: Efficient Large-Scale MoE Training via Resource Modeling and Pipelined Hybrid Parallelism

DGX agent

arXiv:2605.05049v1 Announce Type: cross Abstract: Frontier models increasingly adopt Mixture-of-Experts (MoE) architectures to achieve large-model performance at reduced cost. However, training MoE mo

hardwarearxiv-cs-lg
7 May 2026
Model Releases

Conservative quantum offline model-based optimization

DGX agent

arXiv:2506.19714v2 Announce Type: replace-cross Abstract: Offline model-based optimization (MBO) refers to the task of optimizing a black-box objective function using only a fixed set of prior input-o

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

ELAS: Efficient Pre-Training of Low-Rank Large Language Models via 2:4 Activation Sparsity

DGX agent

arXiv:2605.03667v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved remarkable capabilities, but their immense computational demands during training remain a critical bottleneck

model-releasesarxiv-cs-lg
6 May 2026
Research

Evaluating Reasoning Models for Queries with Presuppositions

DGX agent

arXiv:2605.03050v1 Announce Type: new Abstract: Millions of users turn to AI models for their information needs. It is conceivable that a large number of user queries contain assumptions that may be f

researcharxiv-cs-cl
6 May 2026
Model Releases

🧠 Introducing NeuralBench: a unified, open-source framework to benchmark NeuroAI models. v1.0: 36 EEG tasks, 94 datasets, task-specific + f…

DGX agent

🧠 Introducing NeuralBench: a unified, open-source framework to benchmark NeuroAI models. v1.0: 36 EEG tasks, 94 datasets, task-specific + foundation models. MEG/fMRI ready. MIT-licensed, FAIR's Brain

model-releasesyann-lecun--x
6 May 2026
Model Releases

ISAAC: Auditing Causal Reasoning in Deep Models for Drug-Target Interaction

DGX agent

arXiv:2605.02962v1 Announce Type: new Abstract: Deep learning models for drug--target interaction (DTI) prediction often achieve strong benchmark performance without necessarily relying on mechanistic

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models

DGX agent

arXiv:2605.03821v1 Announce Type: new Abstract: Existing robot video world models are typically trained with low-level objectives such as reconstruction and perceptual similarity, which are poorly ali

model-releasesarxiv-cs-ro
6 May 2026
Model Releases

Adapting Vision-Language Foundation Model for Next Generation Medical Ultrasound Image Analysis

DGX agent

arXiv:2506.08849v4 Announce Type: replace Abstract: Vision-Language Foundation Models (VLFMs) exhibit remarkable generalization, yet their direct application to medical ultrasound is severely hindered

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models

DGX agent

arXiv:2506.09082v5 Announce Type: replace Abstract: The rise of vision foundation models (VFMs) calls for systematic evaluation. A common approach pairs VFMs with large language models (LLMs) as gener

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Focus on the Core: Empowering Diffusion Large Language Models by Self-Contrast

DGX agent

arXiv:2605.01373v1 Announce Type: new Abstract: The iterative denoising paradigm of Diffusion Large Language Models (DLMs) endows them with a distinct advantage in global context modeling. However, cu

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Grounding Synthetic Data Generation With Vision and Language Models

DGX agent

arXiv:2603.09625v2 Announce Type: replace Abstract: Deep learning models benefit from increasing data diversity and volume, motivating synthetic data augmentation to improve existing datasets. However

model-releasesarxiv-cs-cv
5 May 2026
Research

Learning a Stochastic Differential Equation Model of Tropical Cyclone Intensification from Reanalysis and Observational Data

DGX agent

arXiv:2601.08116v2 Announce Type: replace Abstract: Tropical cyclones are dangerous natural hazards, but their hazard is challenging to quantify directly from historical datasets due to limited datase

researcharxiv-cs-lg
5 May 2026
Safety

LVLM-Aided Alignment of Task-Specific Vision Models

DGX agent

arXiv:2512.21985v2 Announce Type: replace Abstract: In high-stakes domains, small task-specific vision models are crucial due to their low computational requirements and the availability of numerous m

safetyarxiv-cs-cv
5 May 2026
Model Releases

MolmoAct2: Action Reasoning Models for Real-world Deployment

DGX agent

arXiv:2605.02881v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models aim to provide a single generalist controller for robots, but today's systems fall short on the criteria that matter

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

Open models should compete on cost and specialization, not frontier benchmarks @natolambert puts it well: the right benchmark is savings in …

DGX agent

Open models should compete on cost and specialization, not frontier benchmarks @natolambert puts it well: the right benchmark is savings in compute and time, especially for repetitive agent tasks deep

model-releasesharrison-chase--x
5 May 2026
Model Releases

OphMAE: Bridging Volumetric and Planar Imaging with a Foundation Model for Adaptive Ophthalmological Diagnosis

DGX agent

arXiv:2605.02714v1 Announce Type: new Abstract: The advent of foundation models has heralded a new era in medical artificial intelligence (AI), enabling the extraction of generalizable representations

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Orthographic Constraint Satisfaction and Human Difficulty Alignment in Large Language Models

DGX agent

arXiv:2511.21086v2 Announce Type: replace Abstract: Large language models must satisfy hard orthographic constraints during controlled text generation, yet systematic cross-family evaluation remains l

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Pretraining A Large Language Model using Distributed GPUs: A Memory-Efficient Decentralized Paradigm

DGX agent

arXiv:2602.11543v2 Announce Type: replace Abstract: Pretraining large language models (LLMs) typically requires centralized clusters with thousands of high-memory GPUs (e.g., H100/A100). Recent decent

model-releasesarxiv-cs-cl
5 May 2026
Research

Consistent Diffusion Language Models

DGX agent

arXiv:2605.00161v1 Announce Type: new Abstract: Diffusion language models (DLMs) are an attractive alternative to autoregressive models because they promise sublinear-time, parallel generation, yet pr

researcharxiv-cs-lg
4 May 2026
Local Ai

RadLite: Multi-Task LoRA Fine-Tuning of Small Language Models for CPU-Deployable Radiology AI

DGX agent

arXiv:2605.00421v1 Announce Type: new Abstract: Large language models (LLMs) show promise in radiology but their deployment is limited by computational requirements that preclude use in resource-const

local-aiarxiv-cs-cl
4 May 2026
Research

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning

DGX agent

arXiv:2603.17837v3 Announce Type: replace-cross Abstract: During conversational interactions, humans subconsciously engage in concurrent thinking while listening to a speaker. Although this internal c

researcharxiv-cs-cl
4 May 2026
Model Releases

CL-bench Life: Can Language Models Learn from Real-Life Context?

DGX agent

arXiv:2604.27043v1 Announce Type: new Abstract: Today's AI assistants such as OpenClaw are designed to handle context effectively, making context learning an increasingly important capability for mode

model-releasesarxiv-cs-cl
1 May 2026
Research

Efficient-DLM: From Autoregressive to Diffusion Language Models, and Beyond in Speed

DGX agent

arXiv:2512.14067v2 Announce Type: replace-cross Abstract: Diffusion language models (dLMs) have emerged as a promising paradigm that enables parallel, non-autoregressive generation, but their learning

researcharxiv-cs-ai
1 May 2026
Model Releases

HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation

DGX agent

arXiv:2604.28196v1 Announce Type: new Abstract: Driving world models serve as a pivotal technology for autonomous driving by simulating environmental dynamics. However, existing approaches predominant

model-releasesarxiv-cs-cv
1 May 2026
Agents

Heterogeneous Scientific Foundation Model Collaboration

DGX agent

arXiv:2604.27351v1 Announce Type: new Abstract: Agentic large language model systems have demonstrated strong capabilities. However, their reliance on language as the universal interface fundamentally

agentsarxiv-cs-ai
1 May 2026
Model Releases

Useless but Safe? Benchmarking Utility Recovery with User Intent Clarification in Multi-Turn Conversations

DGX agent

arXiv:2604.27093v1 Announce Type: cross Abstract: Current LLM safety alignment techniques improve model robustness against adversarial attacks, but overlook whether and how LLMs can recover helpfulnes

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling

DGX agent

arXiv:2604.28185v1 Announce Type: new Abstract: Recent visual generation models have made major progress in photorealism, typography, instruction following, and interactive editing, yet they still str

model-releasesarxiv-cs-cv
1 May 2026
Applications

When Your LLM Reaches End-of-Life: A Framework for Confident Model Migration in Production Systems

DGX agent

arXiv:2604.27082v1 Announce Type: new Abstract: We present a framework for migrating production Large Language Model (LLM) based systems when the underlying model reaches end-of-life or requires repla

applicationsarxiv-cs-ai
1 May 2026
Model Releases

Cross-Domain Transfer of Hyperspectral Foundation Models

DGX agent

arXiv:2604.26478v1 Announce Type: new Abstract: Hyperspectral imaging (HSI) semantic segmentation typically relies on in-domain training, but limited data availability often restricts model performanc

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Goblin and related magical mentions were overrewarded in training, and the behavior was reinforced over successive models. We removed the go…

DGX agent

Goblin and related magical mentions were overrewarded in training, and the behavior was reinforced over successive models. We removed the goblin-affine reward signal for future models, and filtered tr

model-releasesopenai--x
30 Apr 2026
Model Releases

Increasingly, I think, we will see a gap between what you can do with frontier model APIs & what you can do with the native apps from the fr…

DGX agent

Increasingly, I think, we will see a gap between what you can do with frontier model APIs & what you can do with the native apps from the frontier labs (Codex, Claude Code). Models developed and train

model-releasesethan-mollick--x
30 Apr 2026
Model Releases

LLM Psychosis: A Theoretical and Diagnostic Framework for Reality-Boundary Failures in Large Language Models

DGX agent

arXiv:2604.25934v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) as interactive agents has exposed a category of behavioral failure that prevailing terminology, princip

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

OpenAI says its models, starting with GPT-5.1, 'increasingly mentioned goblins, gremlins, and other creatures', leading to prompt instructions to mitigate it (OpenAI)

DGX agent

OpenAI: OpenAI says its models, starting with GPT-5.1, “increasingly mentioned goblins, gremlins, and other creatures”, leading to prompt instructions to mitigate it — Starting with GPT-5.1, our model

model-releasestechmeme
30 Apr 2026
Model Releases

OpenAI’s new security model is for ‘critical cyber defenders’ only

DGX agent

OpenAI is preparing to launch a new frontier cybersecurity model, GPT-5.5-Cyber. CEO Sam Altman said the model will not be available to the general public, but will be first rolled out to a select gro

model-releasesthe-verge-ai
30 Apr 2026
Model Releases

Time Blindness: Why Video-Language Models Can't See What Humans Can?

DGX agent

arXiv:2505.24867v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have made impressive strides in understanding spatio-temporal relationships in videos. Howeve

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

$3/million output tokens. Qwen 3.5 Plus is basically a frontier model. Let that sink in.

DGX agent

$3/million output tokens. Qwen 3.5 Plus is basically a frontier model. Let that sink in. Introducing Qwen3.6-Plus from @Alibaba_Qwen, a 1M-context model built for real-world agents, agentic coding, an

model-releasestogether-ai--x
29 Apr 2026
Safety

Beyond Accuracy: Benchmarking Cross-Task Consistency in Unified Multimodal Models

DGX agent

arXiv:2604.25072v1 Announce Type: new Abstract: Unified Multimodal Models (uMMs) aim to support both visual understanding and visual generation within a shared representation. However, existing evalua

safetyarxiv-cs-cv
29 Apr 2026
Model Releases

Evaluating Computational Pathology Foundation Models for Prostate Cancer Grading under Distribution Shifts

DGX agent

arXiv:2410.06723v2 Announce Type: replace-cross Abstract: Pathology foundation models (PFMs) have emerged as powerful pretrained encoders for computational pathology, but their robustness under clinic

model-releasesarxiv-cs-cv
29 Apr 2026
← Previous
1…6869707172…1259
Next →