AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
12 May 2026

Explicit Reasoning Makes Better Judges: A Systematic Study on Accuracy, Efficiency, and Robustness

Model ReleasesDGX agent

arXiv:2509.13332v2 Announce Type: replace Abstract: As Large Language Models (LLMs) are increasingly adopted as automated judges in benchmarking and reward modeling, ensuring their reliability, effici

Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models

HardwareDGX agent

arXiv:2605.09681v1 Announce Type: new Abstract: Autoregressive (AR) video diffusion models adopt a streaming generation framework, enabling long-horizon video generation with real-time responsiveness,

HairGPT: Strand-as-Language Autoregressive Modeling for Realistic 3D Hairstyle Synthesis

TutorialsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.08824v1 Announce Type: cross Abstract: Hair is a rich medium of visual and cultural expression, yet its digital modeling remains challenging due to the duality of fluidity and structure. Ma

HapticLDM: A Diffusion Model for Text-to-Vibrotactile Generation

SafetyDGX agent

arXiv:2605.09971v1 Announce Type: cross Abstract: Text-to-vibration generation converts natural language into haptic feedback, enabling vibration-effect designers to get scenarios-fitted vibrations mo

Heteroscedastic Diffusion for Multi-Agent Trajectory Modeling

AgentsDGX agent

arXiv:2605.10717v1 Announce Type: cross Abstract: Multi-agent trajectory modeling traditionally focuses on forecasting, often neglecting more general tasks like trajectory completion, which is essenti

How Much Do Circuits Tell Us? Measuring the Consistency and Specificity of Language Model Circuits

ResearchDGX agent

arXiv:2605.08348v1 Announce Type: new Abstract: The circuits framework in mechanistic interpretability aims to identify causally important sparse subgraphs of model components, typically evaluated by

I think frontier model writing is good! It often has a sense of style & tone, variations in sentence structure & length, some great phrasing…

ApplicationsDGX agent

I think frontier model writing is good! It often has a sense of style & tone, variations in sentence structure & length, some great phrasing, etc But it also has some weak spots (fiction!) & clear tic

Improved Mean Flows: On the Challenges of Fastforward Generative Models

ResearchDGX agent

arXiv:2512.02012v2 Announce Type: replace Abstract: MeanFlow (MF) has recently been established as a framework for one-step generative modeling. However, its ``fastforward'' nature introduces key chal

Internalizing Safety Understanding in Large Reasoning Models via Verification

SafetyDGX agent

arXiv:2605.08930v1 Announce Type: new Abstract: While explicit Chain-of-Thought (CoT) empowers large reasoning models (LRMs), it enables the generation of riskier final answers. Current alignment para

Learning Graph Foundation Models on Riemannian Graph-of-Graphs

ResearchDGX agent

arXiv:2605.09993v1 Announce Type: new Abstract: Graph foundation models (GFMs), pretrained on massive graph data, have transformed graph machine learning by supporting general-purpose reasoning across

LLaVA-CKD: Bottom-Up Cascaded Knowledge Distillation for Vision-Language Models

ApplicationsDGX agent

arXiv:2605.10641v1 Announce Type: cross Abstract: Large Vision-Language Models (VLMs) are successful in addressing a multitude of vision-language understanding tasks, such as Visual Question Answering

Measuring Embedding Sensitivity to Authorial Style in French: Comparing Literary Texts with Language Model Rewritings

ResearchDGX agent

arXiv:2605.10606v1 Announce Type: cross Abstract: Large language models (LLMs) can convincingly imitate human writing styles, yet it remains unclear how much stylistic information is encoded in embedd

Metacognitive Behavioral Tuning of Large Language Models for Multi-Hop Question Answering

ResearchDGX agent

arXiv:2602.22508v2 Announce Type: replace Abstract: Large Language Models (LLMs) often produce incorrect answers on multi-hop question answering even when the reasoning trace already contains a correc

Mid-Training with Self-Generated Data Improves Reinforcement Learning in Language Models

SafetyDGX agent

arXiv:2605.08472v1 Announce Type: new Abstract: The effectiveness of Reinforcement Learning (RL) in Large Language Models (LLMs) depends on the nature and diversity of the data used before and during

Mitigating Watermark Forgery in Generative Models via Randomized Key Selection

ResearchDGX agent

arXiv:2507.07871v4 Announce Type: replace-cross Abstract: Watermarking enables GenAI providers to verify whether content was generated by their models. A watermark is a hidden signal in the content, w

Network-Efficient World Model Token Streaming

ResearchDGX agent

arXiv:2605.09886v1 Announce Type: new Abstract: Generative driving world models rely on compact latent state representations that must be efficiently transmitted and synchronized across distributed co

NoiseRater: Meta-Learned Noise Valuation for Diffusion Model Training

ResearchDGX agent

arXiv:2605.08144v1 Announce Type: cross Abstract: Diffusion models have achieved remarkable success across a wide range of generative tasks, yet their training paradigm largely treats injected noise a

Restoration-Aligned Generative Flow Models for Blind Motion Deblurring

ResearchDGX agent

arXiv:2605.08854v1 Announce Type: new Abstract: Generative flow models offer powerful priors learned from large-scale natural images, but directly adapting them to restoration tasks such as motion deb

SDiaReward: Modeling and Benchmarking Spoken Dialogue Rewards with Modality and Colloquialness

Model ReleasesDGX agent

arXiv:2603.14889v2 Announce Type: replace-cross Abstract: The rapid evolution of end-to-end spoken dialogue systems demands transcending mere textual semantics to incorporate paralinguistic nuances an

SenseBench: A Benchmark for Remote Sensing Low-Level Visual Perception and Description in Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.10576v1 Announce Type: cross Abstract: Low-level visual perception underpins reliable remote sensing (RS) image analysis, yet current image quality assessment (IQA) methods output uninterpr

Sequential Causal Discovery with Noisy Language Model Priors

Model ReleasesDGX agent

arXiv:2506.16234v2 Announce Type: replace Abstract: Causal discovery from observational data typically assumes access to complete data and availability of perfect domain experts. In practice, data oft

SLAM: Structural Linguistic Activation Marking for Language Models

Model ReleasesDGX agent

arXiv:2605.05443v2 Announce Type: replace-cross Abstract: LLM watermarks must be detectable without compromising text quality, yet most existing schemes bias the next-token distribution and pay for de

Structure-Centric Graph Foundation Model via Geometric Bases

ResearchDGX agent

arXiv:2605.08689v1 Announce Type: cross Abstract: Graph foundation models (GFMs) seek transferable representations across graph domains but are limited by structural heterogeneity and incompatible nod

Sub-JEPA: Subspace Gaussian Regularization for Stable End-to-End World Models

SafetyDGX agent

arXiv:2605.09241v1 Announce Type: cross Abstract: Joint-Embedding Predictive Architectures (JEPAs) provide a simpleframework for learning world models by predicting future latent representations.Howev

TeleResilienceBench: Quantifying Resilience for LLM Reasoning in Telecommunications

Model ReleasesDGX agent

arXiv:2605.09929v1 Announce Type: new Abstract: Deploying large language models in telecommunications requires more than task accuracy. In realistic workflows, a model may inherit partially completed

Training Reasoning Models on Saturated Problems via Failure-Prefix Conditioning

ResearchDGX agent

arXiv:2601.20829v2 Announce Type: replace-cross Abstract: As Reinforcement Learning with Verifiable Rewards (RLVR) substantially improves the reasoning abilities of large language models (LLMs), a new

Trustworthy AI: Ensuring Reliability and Accountability from Models to Agents

SafetyDGX agent

arXiv:2605.08964v1 Announce Type: new Abstract: In this thesis, we develop algorithms with theoretical guarantees for ensuring reliability and accountability of Machine Learning (ML) systems. As ML sy

Unified Modeling of Lane and Lane Topology for Driving Scene Reasoning

Model ReleasesDGX agent

arXiv:2605.08911v1 Announce Type: new Abstract: Autonomous vehicles need to perceive not only physical elements in the driving scene, such as lane lines and traffic lights, but also logical elements l

V4FinBench: Benchmarking Tabular Foundation Models, LLMs, and Standard Methods on Corporate Bankruptcy Prediction

Model ReleasesDGX agent

arXiv:2605.10896v1 Announce Type: new Abstract: Corporate bankruptcy prediction is a high-stakes financial task characterized by severe class imbalance and multi-horizon forecasting demands. Public da

Virtual Personas for Language Models via an Anthology of Backstories

ResearchDGX agent

arXiv:2407.06576v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are trained from vast repositories of text authored by millions of distinct authors, reflecting an enormous diver

What Will Happen Next: Large Models-Driven Deduction for Emergency Instances

Model ReleasesDGX agent

arXiv:2605.08599v1 Announce Type: new Abstract: Traditional simulation methods reproduce occurred emergency instances through presetting to assist people in risk assessment and emergency decision-maki

11 May 2026

Arrow: A Foundation Model for Causal Discovery

ResearchDGX agent

arXiv:2605.07204v1 Announce Type: new Abstract: We introduce Arrow, a foundation model for zero-shot causal discovery on observational tabular data. Arrow factorizes a directed acyclic graph into an u

Building Blocks for Foundation Model Training and Inference on AWS

ToolsDGX agent

This article discusses AWS infrastructure, tools, and services designed to support the training and deployment of large foundation models, enabling machine learning practitioners to leverage AWS's com

Don't Ignore the Tail: Decoupling top-K Probabilities for Efficient Language Model Distillation

ResearchDGX agent

arXiv:2602.20816v3 Announce Type: replace Abstract: The core learning signal used in language model distillation is the standard Kullback-Leibler (KL) divergence between the student and teacher distri

FinReasoning: A Hierarchical Benchmark for Reliable Financial Research Reporting

Model ReleasesDGX agent

arXiv:2603.19254v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in financial research workflows, where their role is evolving from single-model assistance fo

Flock: A Knowledge Graph Foundation Model via Learning on Random Walks

TutorialsDGX agent

arXiv:2510.01510v3 Announce Type: replace Abstract: We study the problem of zero-shot link prediction on knowledge graphs (KGs), which requires models to generalize to novel entities and novel relatio

From Model to Data (M2D): Shifting Complexity from GNNs to Graphs for Transparent Graph Learning

SafetyDGX agent

arXiv:2605.06814v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) achieve high performance but can be opaque to humans, making it difficult to understand and compare the many proposed archi

Guidance Is Not a Hyperparameter: Learning Dynamic Control in Diffusion Language Models

SafetyDGX agent

arXiv:2605.07701v1 Announce Type: new Abstract: Classifier-Free Guidance (CFG) is a widely used mechanism for controlling diffusion-based generative models, yet its guidance scale is typically treated

Hallucination Detection via Activations of Open-Weight Proxy Analyzers

Model ReleasesDGX agent

arXiv:2605.07209v1 Announce Type: cross Abstract: We introduce a proxy-analyzer framework for detecting hallucinations in large language models. Instead of looking inside the generating model, our sys

Haven’t tried this but it seems very neat… Yet all of the demos (except maybe one) are the model being fun and/or annoying by correcting or …

ApplicationsDGX agent

Haven’t tried this but it seems very neat… Yet all of the demos (except maybe one) are the model being fun and/or annoying by correcting or reminding in real time. There are obvious uses for this sort

Hierarchical Dual-Subspace Decoupling for Continual Learning in Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.07512v1 Announce Type: new Abstract: Class-incremental learning aims to continuously acquire new knowledge while preserving previously learned information, thereby mitigating catastrophic f

InvThink: Premortem Reasoning for Safer Language Models

SafetyDGX agent

arXiv:2510.01569v3 Announce Type: replace Abstract: We present InvThink, a training and prompting framework that requires the model to enumerate, analyze, and constrain potential failures before gener

Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment

SafetyDGX agent

arXiv:2605.08064v1 Announce Type: new Abstract: Spatial intelligence in vision-language models (VLMs) attracts research interest with the practical demand to reason in the 3D world.Despite promising r

Rethinking Dense Sequential Chains: Reasoning Language Models Can Extract Answers from Sparse, Order-Shuffling Chain-of-Thoughts

ResearchDGX agent

arXiv:2605.07307v1 Announce Type: new Abstract: Modern reasoning language models generate dense, sequential chain-of-thought traces implicitly assuming that every token contributes and that steps must

Retina-RAG: Retrieval-Augmented Vision-Language Modeling for Joint Retinal Diagnosis and Clinical Report Generation

Model ReleasesDGX agent

arXiv:2605.06173v2 Announce Type: replace-cross Abstract: Diabetic Retinopathy (DR) is a leading cause of preventable blindness among working-age adults worldwide, yet most automated screening systems

SLOPE: Optimistic Potential Landscape Shaping for Model-based Reinforcement Learning

TutorialsDGX agent

arXiv:2602.03201v3 Announce Type: replace Abstract: Model-based reinforcement learning (MBRL) is sample-efficient but struggles in sparse reward settings. A critical bottleneck arises from the lack of

STARFlow2: Bridging Language Models and Normalizing Flows for Unified Multimodal Generation

ResearchDGX agent

arXiv:2605.08029v1 Announce Type: new Abstract: Deep generative models have advanced rapidly across text and vision, motivating unified multimodal systems that can understand, reason over, and generat

Thinking Machines Lab details interaction models, which can think and respond in real time, letting users and AI interact continuously for better collaboration (Thinking Machines Lab)

IndustryDGX agent

Thinking Machines Lab: Thinking Machines Lab details interaction models, which can think and respond in real time, letting users and AI interact continuously for better collaboration — Today, we're an

Toward Privileged Foundation Models:LUPI for Accelerated and Improved Learning

ResearchDGX agent

arXiv:2605.07799v1 Announce Type: cross Abstract: Training foundation models is computationally intensive and often slow to converge.We introduce PIQL,Privileged Information for Quick and Quality Lear

9 May 2026

What is the best image model for seed variation out of the box?

Local AiDGX agent

This discussion thread examines which image generation models provide the best native seed variation capabilities—the ability to generate diverse images from the same prompt by varying the seed parame

8 May 2026

Improving Bash Generation in Small Language Models with Grammar-Constrained Decoding

HardwareDGX agent

Grammar-constrained decoding modifies language model generation by applying grammar constraints at each step to block structurally invalid tokens , ensuring syntactically correct Bash command generati

7 May 2026

Capacity-Aware Mixture Law Enables Efficient LLM Data Optimization

Model ReleasesDGX agent

arXiv:2603.08022v2 Announce Type: replace Abstract: A data mixture refers to how different data sources are combined to train large language models, and selecting an effective mixture is crucial for o

Concurrence of Symmetry Breaking and Nonlocality Phase Transitions in Diffusion Models

Local AiDGX agent

arXiv:2605.04830v1 Announce Type: new Abstract: Diffusion models undergo a phase transition in a critical time window during generation dynamics, with two complementary diagnoses of criticality. The s

KGLAMP: Knowledge Graph-guided Language model for Adaptive Multi-robot Planning and Replanning

Model ReleasesDGX agent

arXiv:2602.04129v2 Announce Type: replace Abstract: Heterogeneous multi-robot systems are increasingly used in long-horizon missions requiring coordinated planning across diverse capabilities. However

Norm Anchors Make Model Edits Last

ResearchDGX agent

arXiv:2602.02543v3 Announce Type: replace Abstract: Sequential Locate-and-Edit (L&E) model editing can fail abruptly after many edits. We identify and formalize this failure as a positive norm-feedbac

SafeRedir: Prompt Embedding Redirection for Robust Unlearning in Image Generation Models

SafetyDGX agent

arXiv:2601.08623v2 Announce Type: replace Abstract: Image generation models (IGMs), while capable of producing impressive and creative content, often memorize a wide range of undesirable concepts from

SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation

Model ReleasesDGX agent

arXiv:2511.06754v3 Announce Type: replace-cross Abstract: Inspired by how humans reason over discrete objects and their relationships, we explore whether compact object-centric and object-relation rep

Threshold-Guided Optimization for Visual Generative Models

SafetyDGX agent

arXiv:2605.04653v1 Announce Type: new Abstract: Aligning large visual generative models with human feedback is often performed through pairwise preference optimization. While such approaches are conce

Towards Distillation-Resistant Large Language Models: An Information-Theoretic Perspective

ResearchDGX agent

arXiv:2602.03396v3 Announce Type: replace Abstract: Proprietary large language models (LLMs) embody substantial economic value and are generally exposed only as black-box APIs, yet adversaries can sti

6 May 2026

Boosting Team Modeling through Tempo-Relational Representation Learning

ApplicationsDGX agent

arXiv:2507.13305v2 Announce Type: replace Abstract: Team modeling remains a fundamental challenge at the intersection of Artificial Intelligence and Social Sciences. Although a variety of computationa

← Previous
1…161162163164165…1010
Next →