AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,574 results
Tutorials

the best agents aren't just built with the best models: they're built with harnesses purpose-built for the task at hand here's a guide on ho…

DGX agent

the best agents aren't just built with the best models: they're built with harnesses purpose-built for the task at hand here's a guide on how to build a harness that's really good at feeding the model

tutorialsharrison-chase--x
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Industry

theUSshould lead on AI by continuing to develop the very best models, making sure they're safe, and getting cyber tools into the hands of tr…

DGX agent

Sam Altman argues that US leadership in artificial intelligence requires three concurrent priorities: advancing cutting-edge AI model development, ensuring these models incorporate robust safety measu

industrysam-altman--x
3 Jun 2026
Applications

Ultralytics YOLO26: Unified Real-Time End-to-End Vision Models

DGX agent

arXiv:2606.03748v1 Announce Type: cross Abstract: Real-time vision demands models that are accurate, efficient, and simple to deploy across diverse hardware. The YOLO family has become widely deployed

applicationsarxiv-cs-ai
3 Jun 2026
Research

Value-Aware Stochastic KV Cache Eviction for Reasoning Models

DGX agent

arXiv:2606.03928v1 Announce Type: cross Abstract: Reasoning models improve accuracy through extended chains of thought, but their long outputs create a memory and compute bottleneck. KV cache eviction

researcharxiv-cs-cl
3 Jun 2026
Local Ai

A Lightweight Deep Learning-based Model for Ranking Influential Nodes in Complex Networks

DGX agent

arXiv:2507.19702v1 Announce Type: cross Abstract: Identifying influential nodes in complex networks is a critical task with a wide range of applications across different domains. However, existing app

local-aiarxiv-cs-ai
2 Jun 2026
Applications

An Algebraic View of the Expressivity of Recurrent Language Models

DGX agent

arXiv:2606.01765v1 Announce Type: cross Abstract: What formal languages can a recurrent neural language model recognize? Formal results in the literature conflict: some authors report Turing-completen

applicationsarxiv-cs-cl
2 Jun 2026
Applications

Bayesian meta-learning for modeling Alzheimer's disease progression

DGX agent

arXiv:2606.02228v1 Announce Type: cross Abstract: Predicting whether an individual with Alzheimer's disease will experience mild or severe disease progression is essential for personalized treatment.

applicationsarxiv-cs-cv
2 Jun 2026
Research

BERT4beam: Large AI Model Enabled Generalized Beamforming Optimization

DGX agent

arXiv:2509.11056v2 Announce Type: replace-cross Abstract: Artificial intelligence (AI) is anticipated to emerge as a pivotal enabler for the forthcoming sixth-generation (6G) wireless communication sy

researcharxiv-cs-lg
2 Jun 2026
Safety

Breaking the Reversal Curse in Autoregressive Language Models via Identity Bridge

DGX agent

arXiv:2602.02470v2 Announce Type: replace Abstract: Autoregressive large language models (LLMs) have achieved remarkable success in many complex tasks, yet they can still fail in very simple logical r

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

CARTE: A Benchmark for Mapping Language Model Knowledge Across France

DGX agent

arXiv:2606.01995v1 Announce Type: new Abstract: We introduce CARTE 1 (Culturally Anchored Regional-Territorial Evaluation), a multiplechoice benchmark for evaluating the ability of large language mode

model-releasesarxiv-cs-cl
2 Jun 2026
Safety

d2: Improving Reasoning in Diffusion Language Models via Trajectory Likelihood Estimation

DGX agent

arXiv:2509.21474v4 Announce Type: replace Abstract: While diffusion language models (DLMs) have achieved competitive performance in text generation, improving their reasoning ability with reinforcemen

safetyarxiv-cs-lg
2 Jun 2026
Research

Decoding in Order-Agnostic Language Models: Chain-Rule Deviation and Uniform Spreading

DGX agent

arXiv:2606.00997v1 Announce Type: new Abstract: Order-agnostic language models (OALMs), including discrete diffusion language models (dLLMs), are trained to predict masked tokens under arbitrary condi

researcharxiv-cs-cl
2 Jun 2026
Research

Efficient Test-time Inference for Generative Planning Models

DGX agent

arXiv:2606.00618v1 Announce Type: new Abstract: Generative models have emerged as a powerful paradigm for AI planning, yet their performance remains constrained by the training data distribution. One

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Error Bounds for a Diffusion Model-Based Drift Estimator

DGX agent

arXiv:2606.02115v1 Announce Type: cross Abstract: Parameter estimation in stochastic differential equations is a classical statistical problem of much importance in many scientific fields. Recent work

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Evaluating Real-World Generalizability of Algorithm Selection Models

DGX agent

arXiv:2606.02016v1 Announce Type: new Abstract: Algorithm Selection (AS) aims to automatically identify the most suitable optimization algorithm for a given problem instance by leveraging measurable p

model-releasesarxiv-cs-lg
2 Jun 2026
Research

Evaluating the Performance of Deep Learning Models in Whole-body Dynamic 3D Posture Prediction During Load-reaching Activities

DGX agent

arXiv:2511.20615v2 Announce Type: replace-cross Abstract: This study aimed to explore the application of deep neural networks for whole-body human posture prediction during dynamic load-reaching activ

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Exploring the Capabilities of Large Language Model Encoders for Image-Text Retrieval in Chest X-rays

DGX agent

arXiv:2509.15234v2 Announce Type: replace Abstract: Multimodal learning from paired medical images and clinical text is a central challenge in medical data-driven informatics, where effective cross-mo

model-releasesarxiv-cs-cv
2 Jun 2026
Hardware

Eyettention II: A Dual-Sequence Architecture for Modeling Fixation Location, Within-Word Landing Position, and Fixation Duration in Reading

DGX agent

arXiv:2606.01964v1 Announce Type: new Abstract: The way our eyes move while reading provides valuable insights into both the reader's cognitive processes and the properties of the text. In particular,

hardwarearxiv-cs-cl
2 Jun 2026
Safety

Failure of contextual invariance in large language models

DGX agent

arXiv:2603.23485v2 Announce Type: replace-cross Abstract: Standard evaluation practices assume that large language model (LLM) outputs are stable when prompts are embedded in contextually equivalent d

safetyarxiv-cs-ai
2 Jun 2026
Research

FATE-VLA:Failue-aware test generation for vision-language-action models

DGX agent

arXiv:2606.02307v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are increasingly used as generalist robot policies, yet their evaluation still relies largely on static benchmarks t

researcharxiv-cs-ro
2 Jun 2026
Model Releases

Fine-Tuning Diffusion Models for Molecular Generation via Reinforcement Learning and Fast Sampling

DGX agent

arXiv:2606.01220v1 Announce Type: cross Abstract: Generating molecules that simultaneously satisfy drug-like properties and conform to the 3D structure of a target protein is a core challenge in struc

model-releasesarxiv-cs-ai
2 Jun 2026
Applications

Fine-Tuning Without Forgetting In-Context Learning: A Theoretical Analysis of Linear Attention Models

DGX agent

arXiv:2602.23197v2 Announce Type: replace Abstract: Transformer-based large language models exhibit in-context learning, enabling adaptation to downstream tasks via few-shot prompting with demonstrati

applicationsarxiv-cs-cl
2 Jun 2026
Hardware

FLARE: Diffusion for Hybrid Language Model

DGX agent

arXiv:2606.01774v1 Announce Type: cross Abstract: Autoregressive (AR) large language models (LLMs) have achieved broad practical success, but sequential decoding remains a key bottleneck for low-laten

hardwarearxiv-cs-ai
2 Jun 2026
Safety

From Demonstrations to Rewards: Test-Time Prompt Optimization for VLM Reward Models

DGX agent

arXiv:2606.00083v1 Announce Type: cross Abstract: Reinforcement learning relies on accurate reward functions, which are often hand-crafted or even unavailable in real-world applications, such as robot

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

How AI Fails: An Interactive Pedagogical Tool for Demonstrating Dialectal Bias in Automated Toxicity Models

DGX agent

arXiv:2511.06676v3 Announce Type: replace Abstract: Now that AI-driven moderation has become pervasive in everyday life, we often hear claims that 'the AI is biased'. While this is often said jokingly

model-releasesarxiv-cs-cl
2 Jun 2026
Research

HumanNOVA: Photorealistic, Universal and Rapid 3D Human Avatar Modeling from a Single Image

DGX agent

arXiv:2606.02573v1 Announce Type: new Abstract: In this paper, we present HumanNOVA, a photorealistic, universal, and rapid model for generating 3D human avatars from a single RGB image. Achieving bot

researcharxiv-cs-cv
2 Jun 2026
Research

Intercepting the Future: Latent-Space Predictive World Model for Dynamic VLA Manipulation

DGX agent

arXiv:2606.02486v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models generalize across static manipulation but fail when objects move during task execution. They map the current observa

researcharxiv-cs-ro
2 Jun 2026
Research

Lessons from the Trenches on Reproducible Evaluation of Language Models

DGX agent

arXiv:2405.14782v3 Announce Type: replace Abstract: Reliable evaluation of language models (LMs) remains an open challenge. Re- searchers and engineers face methodological issues such as the sensitivi

researcharxiv-cs-cl
2 Jun 2026
Research

Limits of Spatial Imagery Reasoning in Frontier LLM Models

DGX agent

arXiv:2603.26779v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated impressive reasoning capabilities, yet they struggle with spatial tasks that require mental sim

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Physics-Encoded Inverse Modeling for Arctic Snow Depth Prediction

DGX agent

arXiv:2601.17074v4 Announce Type: replace-cross Abstract: Accurate estimation in time-varying inverse problems under limited and sparse observations remains a fundamental challenge across scientific d

model-releasesarxiv-cs-ai
2 Jun 2026
Local Ai

Physics-Informed Modeling and Control of Emergent Behaviors in Robot Swarms

DGX agent

arXiv:2606.01597v1 Announce Type: new Abstract: Robot swarms can exhibit coherent collective behaviors through local perception, limited communication and decentralized decision-making, yet modeling a

local-aiarxiv-cs-ro
2 Jun 2026
Model Releases

Plan-R1: Safe and Feasible Trajectory Planning as Language Modeling

DGX agent

arXiv:2505.17659v4 Announce Type: replace-cross Abstract: Safe and feasible trajectory planning is critical for real-world autonomous driving systems. However, existing learning-based planners rely he

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

Policy and World Modeling Co-Training for Language Agents

DGX agent

arXiv:2606.02388v1 Announce Type: cross Abstract: Reinforcement learning (RL) improves large language model (LLM) agents by teaching them which actions lead to high rewards, but provides little superv

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

PUMA: Layer-Pruned Language Model for Efficient Unified Multimodal Retrieval with Modality-Adaptive Learning

DGX agent

arXiv:2507.08064v3 Announce Type: replace-cross Abstract: As multimedia content expands, the demand for unified multimodal retrieval (UMR) in real-world applications increases. Recent work leverages m

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

RAIGen: Rare Attribute Identification in Text-to-Image Generative Models

DGX agent

arXiv:2602.06806v2 Announce Type: replace Abstract: Text-to-image diffusion models achieve impressive generation quality but inherit and amplify training-data biases, skewing coverage of semantic attr

safetyarxiv-cs-cv
2 Jun 2026
Research

ResMerge: Residual-based Spectral Merging of Large Language Models

DGX agent

arXiv:2606.02252v1 Announce Type: new Abstract: Model merging offers a training-free way to combine multiple post-trained expert models, but merging experts obtained through reinforcement learning (RL

researcharxiv-cs-cl
2 Jun 2026
Model Releases

Retrieval-aligned Tabular Foundation Models Enable Robust Clinical Risk Prediction in Electronic Health Records Under Real-world Constraints

DGX agent

arXiv:2604.01841v2 Announce Type: replace Abstract: Clinical prediction from structured electronic health records (EHRs) is challenging due to high dimensionality, heterogeneity, class imbalance, and

model-releasesarxiv-cs-ai
2 Jun 2026
Local Ai

RuleEdit: Failure-Guided Human-AI Model Editing with Prospective Impact Preview

DGX agent

arXiv:2606.00011v1 Announce Type: cross Abstract: Despite the promise of AI to assist complex decisions, practitioners still lack ways to detect likely failures and inspect the consequences of model e

local-aiarxiv-cs-ai
2 Jun 2026
Model Releases

See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation

DGX agent

arXiv:2603.09292v2 Announce Type: replace-cross Abstract: Measurement of task progress through explicit, actionable milestones is critical for robust robotic manipulation. This progress awareness enab

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Single-Channel Tissue Segmentation via Cross-Modal Distillation from Foundation Models

DGX agent

arXiv:2606.00928v1 Announce Type: new Abstract: Multiplexed fluorescence microscopy improves tissue segmentation by providing complementary channels including nuclear (DAPI) and membrane (E-cadherin),

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement

DGX agent

arXiv:2606.00267v1 Announce Type: cross Abstract: Video world models (WMs) have shown promise for policy evaluation and improvement by imagining realistic future observations conditioned on ego-robot

safetyarxiv-cs-ai
2 Jun 2026
Local Ai

Today we're announcing that hybrid agentic inference is coming to Perplexity Computer. Computer can split tasks between a local model runnin…

DGX agent

Today we're announcing that hybrid agentic inference is coming to Perplexity Computer. Computer can split tasks between a local model running on your machine and frontier models in the cloud. This kee

local-aiperplexity--x
2 Jun 2026
Research

Understanding the Effects of Distractors on Reasoning Vision-Language Models

DGX agent

arXiv:2511.21397v2 Announce Type: replace-cross Abstract: How does irrelevant information (i.e., distractors) affect test-time scaling in vision-language models (VLMs)? Prior work on text-only languag

researcharxiv-cs-ai
2 Jun 2026
Model Releases

VESTA: Visual Exploration with Statistical Tool Agents

DGX agent

arXiv:2606.00384v1 Announce Type: new Abstract: Fitting quantitative models to data is a central step in scientific workflows, yet it remains one of the least automated. Recent agent-based systems lev

model-releasesarxiv-cs-ai
2 Jun 2026
Research

AR Forcing: Towards Long-Horizon Robot Navigation World Model

DGX agent

arXiv:2605.31314v1 Announce Type: new Abstract: The diffusion based robot navigation world models are typically trained using parallel supervision, while autoregressive inference is employed during pa

researcharxiv-cs-ro
1 Jun 2026
Model Releases

Chain-of-Thought Reasoning In The Wild Is Not Always Faithful

DGX agent

arXiv:2503.08679v5 Announce Type: replace Abstract: Recent studies indicate that when faced with explicit biases in prompts, models often omit mentioning these biases in their Chain-of-Thought (CoT) o

model-releasesarxiv-cs-ai
1 Jun 2026
Safety

Configurable Reward Model for Balanced Safety Alignment

DGX agent

arXiv:2605.30487v1 Announce Type: new Abstract: Aligning large language models (LLMs) to heterogeneous and rapidly evolving safety requirements remains a critical challenge. Existing instruction-tuned

safetyarxiv-cs-cl
1 Jun 2026
Safety

Does Visual Information Play a Decisive Role in Vision-Language-Action Model Driving Behavior?

DGX agent

arXiv:2605.31041v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated promising capability in autonomous driving, highlighting the potential of unified multimodal arc

safetyarxiv-cs-ai
1 Jun 2026
← Previous
1…166167168169170…1262
Next →