AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,038 results
Research

Improved Mean Flows: On the Challenges of Fastforward Generative Models

DGX agent

arXiv:2512.02012v2 Announce Type: replace Abstract: MeanFlow (MF) has recently been established as a framework for one-step generative modeling. However, its ``fastforward'' nature introduces key chal

researcharxiv-cs-cv
12 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Internalizing Safety Understanding in Large Reasoning Models via Verification

DGX agent

arXiv:2605.08930v1 Announce Type: new Abstract: While explicit Chain-of-Thought (CoT) empowers large reasoning models (LRMs), it enables the generation of riskier final answers. Current alignment para

safetyarxiv-cs-ai
12 May 2026
Research

Learning Graph Foundation Models on Riemannian Graph-of-Graphs

DGX agent

arXiv:2605.09993v1 Announce Type: new Abstract: Graph foundation models (GFMs), pretrained on massive graph data, have transformed graph machine learning by supporting general-purpose reasoning across

researcharxiv-cs-lg
12 May 2026
Applications

LLaVA-CKD: Bottom-Up Cascaded Knowledge Distillation for Vision-Language Models

DGX agent

arXiv:2605.10641v1 Announce Type: cross Abstract: Large Vision-Language Models (VLMs) are successful in addressing a multitude of vision-language understanding tasks, such as Visual Question Answering

applicationsarxiv-cs-ai
12 May 2026
Research

Measuring Embedding Sensitivity to Authorial Style in French: Comparing Literary Texts with Language Model Rewritings

DGX agent

arXiv:2605.10606v1 Announce Type: cross Abstract: Large language models (LLMs) can convincingly imitate human writing styles, yet it remains unclear how much stylistic information is encoded in embedd

researcharxiv-cs-ai
12 May 2026
Research

Metacognitive Behavioral Tuning of Large Language Models for Multi-Hop Question Answering

DGX agent

arXiv:2602.22508v2 Announce Type: replace Abstract: Large Language Models (LLMs) often produce incorrect answers on multi-hop question answering even when the reasoning trace already contains a correc

researcharxiv-cs-ai
12 May 2026
Safety

Mid-Training with Self-Generated Data Improves Reinforcement Learning in Language Models

DGX agent

arXiv:2605.08472v1 Announce Type: new Abstract: The effectiveness of Reinforcement Learning (RL) in Large Language Models (LLMs) depends on the nature and diversity of the data used before and during

safetyarxiv-cs-ai
12 May 2026
Research

Mitigating Watermark Forgery in Generative Models via Randomized Key Selection

DGX agent

arXiv:2507.07871v4 Announce Type: replace-cross Abstract: Watermarking enables GenAI providers to verify whether content was generated by their models. A watermark is a hidden signal in the content, w

researcharxiv-cs-ai
12 May 2026
Research

Network-Efficient World Model Token Streaming

DGX agent

arXiv:2605.09886v1 Announce Type: new Abstract: Generative driving world models rely on compact latent state representations that must be efficiently transmitted and synchronized across distributed co

researcharxiv-cs-ro
12 May 2026
Research

NoiseRater: Meta-Learned Noise Valuation for Diffusion Model Training

DGX agent

arXiv:2605.08144v1 Announce Type: cross Abstract: Diffusion models have achieved remarkable success across a wide range of generative tasks, yet their training paradigm largely treats injected noise a

researcharxiv-cs-ai
12 May 2026
Research

Restoration-Aligned Generative Flow Models for Blind Motion Deblurring

DGX agent

arXiv:2605.08854v1 Announce Type: new Abstract: Generative flow models offer powerful priors learned from large-scale natural images, but directly adapting them to restoration tasks such as motion deb

researcharxiv-cs-cv
12 May 2026
Model Releases

SDiaReward: Modeling and Benchmarking Spoken Dialogue Rewards with Modality and Colloquialness

DGX agent

arXiv:2603.14889v2 Announce Type: replace-cross Abstract: The rapid evolution of end-to-end spoken dialogue systems demands transcending mere textual semantics to incorporate paralinguistic nuances an

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

SenseBench: A Benchmark for Remote Sensing Low-Level Visual Perception and Description in Large Vision-Language Models

DGX agent

arXiv:2605.10576v1 Announce Type: cross Abstract: Low-level visual perception underpins reliable remote sensing (RS) image analysis, yet current image quality assessment (IQA) methods output uninterpr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Sequential Causal Discovery with Noisy Language Model Priors

DGX agent

arXiv:2506.16234v2 Announce Type: replace Abstract: Causal discovery from observational data typically assumes access to complete data and availability of perfect domain experts. In practice, data oft

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

SLAM: Structural Linguistic Activation Marking for Language Models

DGX agent

arXiv:2605.05443v2 Announce Type: replace-cross Abstract: LLM watermarks must be detectable without compromising text quality, yet most existing schemes bias the next-token distribution and pay for de

model-releasesarxiv-cs-ai
12 May 2026
Research

Structure-Centric Graph Foundation Model via Geometric Bases

DGX agent

arXiv:2605.08689v1 Announce Type: cross Abstract: Graph foundation models (GFMs) seek transferable representations across graph domains but are limited by structural heterogeneity and incompatible nod

researcharxiv-cs-ai
12 May 2026
Safety

Sub-JEPA: Subspace Gaussian Regularization for Stable End-to-End World Models

DGX agent

arXiv:2605.09241v1 Announce Type: cross Abstract: Joint-Embedding Predictive Architectures (JEPAs) provide a simpleframework for learning world models by predicting future latent representations.Howev

safetyarxiv-cs-ai
12 May 2026
Model Releases

TeleResilienceBench: Quantifying Resilience for LLM Reasoning in Telecommunications

DGX agent

arXiv:2605.09929v1 Announce Type: new Abstract: Deploying large language models in telecommunications requires more than task accuracy. In realistic workflows, a model may inherit partially completed

model-releasesarxiv-cs-lg
12 May 2026
Research

Training Reasoning Models on Saturated Problems via Failure-Prefix Conditioning

DGX agent

arXiv:2601.20829v2 Announce Type: replace-cross Abstract: As Reinforcement Learning with Verifiable Rewards (RLVR) substantially improves the reasoning abilities of large language models (LLMs), a new

researcharxiv-cs-ai
12 May 2026
Safety

Trustworthy AI: Ensuring Reliability and Accountability from Models to Agents

DGX agent

arXiv:2605.08964v1 Announce Type: new Abstract: In this thesis, we develop algorithms with theoretical guarantees for ensuring reliability and accountability of Machine Learning (ML) systems. As ML sy

safetyarxiv-cs-lg
12 May 2026
Model Releases

Unified Modeling of Lane and Lane Topology for Driving Scene Reasoning

DGX agent

arXiv:2605.08911v1 Announce Type: new Abstract: Autonomous vehicles need to perceive not only physical elements in the driving scene, such as lane lines and traffic lights, but also logical elements l

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

V4FinBench: Benchmarking Tabular Foundation Models, LLMs, and Standard Methods on Corporate Bankruptcy Prediction

DGX agent

arXiv:2605.10896v1 Announce Type: new Abstract: Corporate bankruptcy prediction is a high-stakes financial task characterized by severe class imbalance and multi-horizon forecasting demands. Public da

model-releasesarxiv-cs-lg
12 May 2026
Research

Virtual Personas for Language Models via an Anthology of Backstories

DGX agent

arXiv:2407.06576v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are trained from vast repositories of text authored by millions of distinct authors, reflecting an enormous diver

researcharxiv-cs-ai
12 May 2026
Model Releases

What Will Happen Next: Large Models-Driven Deduction for Emergency Instances

DGX agent

arXiv:2605.08599v1 Announce Type: new Abstract: Traditional simulation methods reproduce occurred emergency instances through presetting to assist people in risk assessment and emergency decision-maki

model-releasesarxiv-cs-ai
12 May 2026
Research

Arrow: A Foundation Model for Causal Discovery

DGX agent

arXiv:2605.07204v1 Announce Type: new Abstract: We introduce Arrow, a foundation model for zero-shot causal discovery on observational tabular data. Arrow factorizes a directed acyclic graph into an u

researcharxiv-cs-lg
11 May 2026
Tools

Building Blocks for Foundation Model Training and Inference on AWS

DGX agent

This article discusses AWS infrastructure, tools, and services designed to support the training and deployment of large foundation models, enabling machine learning practitioners to leverage AWS's com

toolshugging-face
11 May 2026
Research

Don't Ignore the Tail: Decoupling top-K Probabilities for Efficient Language Model Distillation

DGX agent

arXiv:2602.20816v3 Announce Type: replace Abstract: The core learning signal used in language model distillation is the standard Kullback-Leibler (KL) divergence between the student and teacher distri

researcharxiv-cs-cl
11 May 2026
Model Releases

FinReasoning: A Hierarchical Benchmark for Reliable Financial Research Reporting

DGX agent

arXiv:2603.19254v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in financial research workflows, where their role is evolving from single-model assistance fo

model-releasesarxiv-cs-cl
11 May 2026
Tutorials

Flock: A Knowledge Graph Foundation Model via Learning on Random Walks

DGX agent

arXiv:2510.01510v3 Announce Type: replace Abstract: We study the problem of zero-shot link prediction on knowledge graphs (KGs), which requires models to generalize to novel entities and novel relatio

tutorialsarxiv-cs-lg
11 May 2026
Safety

From Model to Data (M2D): Shifting Complexity from GNNs to Graphs for Transparent Graph Learning

DGX agent

arXiv:2605.06814v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) achieve high performance but can be opaque to humans, making it difficult to understand and compare the many proposed archi

safetyarxiv-cs-lg
11 May 2026
Safety

Guidance Is Not a Hyperparameter: Learning Dynamic Control in Diffusion Language Models

DGX agent

arXiv:2605.07701v1 Announce Type: new Abstract: Classifier-Free Guidance (CFG) is a widely used mechanism for controlling diffusion-based generative models, yet its guidance scale is typically treated

safetyarxiv-cs-cl
11 May 2026
Model Releases

Hallucination Detection via Activations of Open-Weight Proxy Analyzers

DGX agent

arXiv:2605.07209v1 Announce Type: cross Abstract: We introduce a proxy-analyzer framework for detecting hallucinations in large language models. Instead of looking inside the generating model, our sys

model-releasesarxiv-cs-ai
11 May 2026
Applications

Haven’t tried this but it seems very neat… Yet all of the demos (except maybe one) are the model being fun and/or annoying by correcting or …

DGX agent

Haven’t tried this but it seems very neat… Yet all of the demos (except maybe one) are the model being fun and/or annoying by correcting or reminding in real time. There are obvious uses for this sort

applicationsethan-mollick--x
11 May 2026
Model Releases

Hierarchical Dual-Subspace Decoupling for Continual Learning in Vision-Language Models

DGX agent

arXiv:2605.07512v1 Announce Type: new Abstract: Class-incremental learning aims to continuously acquire new knowledge while preserving previously learned information, thereby mitigating catastrophic f

model-releasesarxiv-cs-cv
11 May 2026
Safety

InvThink: Premortem Reasoning for Safer Language Models

DGX agent

arXiv:2510.01569v3 Announce Type: replace Abstract: We present InvThink, a training and prompting framework that requires the model to enumerate, analyze, and constrain potential failures before gener

safetyarxiv-cs-ai
11 May 2026
Safety

Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment

DGX agent

arXiv:2605.08064v1 Announce Type: new Abstract: Spatial intelligence in vision-language models (VLMs) attracts research interest with the practical demand to reason in the 3D world.Despite promising r

safetyarxiv-cs-cv
11 May 2026
Research

Rethinking Dense Sequential Chains: Reasoning Language Models Can Extract Answers from Sparse, Order-Shuffling Chain-of-Thoughts

DGX agent

arXiv:2605.07307v1 Announce Type: new Abstract: Modern reasoning language models generate dense, sequential chain-of-thought traces implicitly assuming that every token contributes and that steps must

researcharxiv-cs-cl
11 May 2026
Model Releases

Retina-RAG: Retrieval-Augmented Vision-Language Modeling for Joint Retinal Diagnosis and Clinical Report Generation

DGX agent

arXiv:2605.06173v2 Announce Type: replace-cross Abstract: Diabetic Retinopathy (DR) is a leading cause of preventable blindness among working-age adults worldwide, yet most automated screening systems

model-releasesarxiv-cs-ai
11 May 2026
Tutorials

SLOPE: Optimistic Potential Landscape Shaping for Model-based Reinforcement Learning

DGX agent

arXiv:2602.03201v3 Announce Type: replace Abstract: Model-based reinforcement learning (MBRL) is sample-efficient but struggles in sparse reward settings. A critical bottleneck arises from the lack of

tutorialsarxiv-cs-lg
11 May 2026
Research

STARFlow2: Bridging Language Models and Normalizing Flows for Unified Multimodal Generation

DGX agent

arXiv:2605.08029v1 Announce Type: new Abstract: Deep generative models have advanced rapidly across text and vision, motivating unified multimodal systems that can understand, reason over, and generat

researcharxiv-cs-cv
11 May 2026
Industry

Thinking Machines Lab details interaction models, which can think and respond in real time, letting users and AI interact continuously for better collaboration (Thinking Machines Lab)

DGX agent

Thinking Machines Lab: Thinking Machines Lab details interaction models, which can think and respond in real time, letting users and AI interact continuously for better collaboration — Today, we're an

industrytechmeme
11 May 2026
Research

Toward Privileged Foundation Models:LUPI for Accelerated and Improved Learning

DGX agent

arXiv:2605.07799v1 Announce Type: cross Abstract: Training foundation models is computationally intensive and often slow to converge.We introduce PIQL,Privileged Information for Quick and Quality Lear

researcharxiv-cs-ai
11 May 2026
Local Ai

What is the best image model for seed variation out of the box?

DGX agent

This discussion thread examines which image generation models provide the best native seed variation capabilities—the ability to generate diverse images from the same prompt by varying the seed parame

local-air-stablediffusion
9 May 2026
Hardware

Improving Bash Generation in Small Language Models with Grammar-Constrained Decoding

DGX agent

Grammar-constrained decoding modifies language model generation by applying grammar constraints at each step to block structurally invalid tokens , ensuring syntactically correct Bash command generati

hardwarenvidia-developer
8 May 2026
Model Releases

Capacity-Aware Mixture Law Enables Efficient LLM Data Optimization

DGX agent

arXiv:2603.08022v2 Announce Type: replace Abstract: A data mixture refers to how different data sources are combined to train large language models, and selecting an effective mixture is crucial for o

model-releasesarxiv-cs-lg
7 May 2026
Local Ai

Concurrence of Symmetry Breaking and Nonlocality Phase Transitions in Diffusion Models

DGX agent

arXiv:2605.04830v1 Announce Type: new Abstract: Diffusion models undergo a phase transition in a critical time window during generation dynamics, with two complementary diagnoses of criticality. The s

local-aiarxiv-cs-lg
7 May 2026
Model Releases

KGLAMP: Knowledge Graph-guided Language model for Adaptive Multi-robot Planning and Replanning

DGX agent

arXiv:2602.04129v2 Announce Type: replace Abstract: Heterogeneous multi-robot systems are increasingly used in long-horizon missions requiring coordinated planning across diverse capabilities. However

model-releasesarxiv-cs-ro
7 May 2026
Research

Norm Anchors Make Model Edits Last

DGX agent

arXiv:2602.02543v3 Announce Type: replace Abstract: Sequential Locate-and-Edit (L&E) model editing can fail abruptly after many edits. We identify and formalize this failure as a positive norm-feedbac

researcharxiv-cs-lg
7 May 2026
← Previous
1…207208209210211…1293
Next →