AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,811 results
Research

A Model of Multi-turn Human Persuadability Using Probabilistic Belief Tracing

DGX agent

arXiv:2606.05330v1 Announce Type: new Abstract: Large language models can shift human beliefs across high-stakes domains, but most persuasion studies rely on pre/post belief change. These endpoint mea

researcharxiv-cs-cl
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Adversarial Attacks Already Tell the Answer: Directional Bias-Guided Test-time Defense for Vision-Language Models

DGX agent

arXiv:2606.06186v1 Announce Type: new Abstract: Vision-Language Models (VLMs), such as CLIP, have shown strong zero-shot generalization but remain highly vulnerable to adversarial perturbations, posin

safetyarxiv-cs-cv
5 Jun 2026
Research

Amortized Nonlinear Model Predictive Control

DGX agent

arXiv:2606.05840v1 Announce Type: cross Abstract: Nonlinear Model Predictive Control requires solving a constrained nonlinear program (NLP) in real-time at every sampling instant, a computational bott

researcharxiv-cs-ro
5 Jun 2026
Tutorials

Can Language Models Learn to Listen?

DGX agent

arXiv:2308.10897v2 Announce Type: replace Abstract: We present a framework for generating appropriate facial responses from a listener in dyadic social interactions based on the speaker's words. Given

tutorialsarxiv-cs-cv
5 Jun 2026
Tutorials

Diff-CA: Separating Common and Salient Factors with Diffusion Models

DGX agent

arXiv:2606.06120v1 Announce Type: new Abstract: Contrastive Analysis aims to separate factors that are common between two data distributions from those that are salient to only one of them. Existing c

tutorialsarxiv-cs-cv
5 Jun 2026
Safety

Epistemic Injustice in Language Models: An Audit of Pretraining Filters and Guardrails

DGX agent

arXiv:2606.05936v1 Announce Type: new Abstract: Modern language models rely on pretraining filters to remove undesirable content from training corpora and inference-time guardrails to suppress undesir

safetyarxiv-cs-cl
5 Jun 2026
Hardware

Flash-WAM: Modality-Aware Distillation for World Action Models

DGX agent

arXiv:2606.05254v1 Announce Type: cross Abstract: World-action models (WAMs) jointly generate future video and robot actions through iterative diffusion, achieving strong performance on manipulation b

hardwarearxiv-cs-cv
5 Jun 2026
Tutorials

Learning Self-Correction in Vision-Language Models via Rollout Augmentation

DGX agent

arXiv:2602.08503v2 Announce Type: replace-cross Abstract: Self-correction is essential for solving complex reasoning problems in vision-language models (VLMs). However, existing reinforcement learning

tutorialsarxiv-cs-cl
5 Jun 2026
Research

Leveraging Large Language Models for Generating Research Topic Ontologies: A Multi-Disciplinary Study

DGX agent

arXiv:2508.20693v2 Announce Type: replace-cross Abstract: Ontologies and taxonomies of research fields are critical for managing and organising scientific knowledge, as they facilitate efficient class

researcharxiv-cs-cl
5 Jun 2026
Research

MS-DKC: A Dataset Knowledge Card Framework for Designing and Adapting Medical Image Segmentation Models

DGX agent

arXiv:2606.06103v1 Announce Type: new Abstract: Medical image segmentation is often framed as a search for stronger architectures, but this can obscure a more fundamental question: what does the datas

researcharxiv-cs-cv
5 Jun 2026
Safety

ReCache: Learning Budget-Aware Caching Schedules for Diffusion Models via REINFORCE

DGX agent

arXiv:2606.06060v1 Announce Type: new Abstract: Modern diffusion models generate high-quality images and videos, but their iterative denoising process makes inference expensive. Feature caching accele

safetyarxiv-cs-cv
5 Jun 2026
Local Ai

Temporal Preference Concepts and their Functions in a Large Language Model

DGX agent

arXiv:2606.05194v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly being deployed to make decisions that require trading off near-term gains against long-term consequences

local-aiarxiv-cs-cl
5 Jun 2026
Research

A Systematic Analysis of Linguistic Features in AI-Generated Text Detection Across Domains and Models

DGX agent

arXiv:2606.04177v1 Announce Type: cross Abstract: Interpretable linguistic features offer a promising approach for explaining why a given text appears machine-generated, particularly for non-expert us

researcharxiv-cs-ai
4 Jun 2026
Safety

Activation Steering of Video Generation Models via Reduced-Order Linear Optimal Control

DGX agent

arXiv:2606.04775v1 Announce Type: cross Abstract: Text-to-video (T2V) models trained on large-scale web data can generate undesired content, motivating interventions that reduce harmful outputs withou

safetyarxiv-cs-ai
4 Jun 2026
Safety

Beyond Symmetric Alignment: Spectral Diagnostics of Modality Imbalance in Vision-Language Models in the Medical Domain

DGX agent

arXiv:2606.04613v1 Announce Type: new Abstract: Vision-Language Models (VLMs) struggle when applied to medical image-text data, yet the tools available to diagnose this failure remain limited. Existin

safetyarxiv-cs-cv
4 Jun 2026
Research

Bounded Hyperbolic Tangent: A Stable and Efficient Alternative to Pre-Layer Normalization in Large Language Models

DGX agent

arXiv:2601.09719v3 Announce Type: replace-cross Abstract: Pre-Layer Normalization (Pre-LN) is the de facto choice for large language models (LLMs) and is crucial for stable pretraining and effective t

researcharxiv-cs-ai
4 Jun 2026
Safety

Culturally Grounded Personas in Large Language Models: Characterization and Alignment with Socio-Psychological Value Frameworks

DGX agent

arXiv:2601.22396v2 Announce Type: replace-cross Abstract: Despite the growing utility of Large Language Models (LLMs) for simulating human behavior, the extent to which these synthetic personas accura

safetyarxiv-cs-ai
4 Jun 2026
Research

Dynamic Infilling Anchors for Format-Constrained Generation in Diffusion Large Language Models

DGX agent

arXiv:2606.04535v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) offer bidirectional attention and parallel generation, enabling them to exploit global context and naturally s

researcharxiv-cs-ai
4 Jun 2026
Safety

Geometry-Aware Distillation for Prompt Tuning Biomedical Vision-Language Models

DGX agent

arXiv:2606.04922v1 Announce Type: cross Abstract: Current prompt-based and adapter-based tuning of vision-language models (VLMs) is attractive for medical imaging, where clinical data sensitivity favo

safetyarxiv-cs-ai
4 Jun 2026
Tutorials

GeoMin: Data-Efficient Semi-Supervised RLVR via Geometric Distribution Modeling

DGX agent

arXiv:2606.04516v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) significantly advances LLM reasoning, yet it faces a dilemma: standard supervised scaling is thr

tutorialsarxiv-cs-ai
4 Jun 2026
Research

How Users Understand Robot Foundation Model Performance through Task Success Rates and Beyond

DGX agent

arXiv:2602.03920v2 Announce Type: replace Abstract: Robot Foundation Models (RFMs) represent a promising approach to developing general-purpose home robots. Given the broad capabilities of RFMs, users

researcharxiv-cs-ro
4 Jun 2026
Local Ai

MAD: Mapping-Aware World Models for Agile Quadrotor Flight

DGX agent

arXiv:2606.04534v1 Announce Type: new Abstract: Agile quadrotor flight in cluttered scenes requires more than a reactive mapping from a depth image to a control command: the vehicle must remember whic

local-aiarxiv-cs-ro
4 Jun 2026
Safety

MaskForge: Structure-Aware Adaptive Attacks for Jailbreaking Diffusion Large Language Models

DGX agent

arXiv:2606.04027v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising partially masked sequences under bidirectional context, exposing a safe

safetyarxiv-cs-ai
4 Jun 2026
Research

Read the Trace, Steer the Path: Trajectory-Aware Reinforcement Learning for Diffusion Language Models

DGX agent

arXiv:2606.04396v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) generate responses by iteratively unmasking and revising many positions in parallel. This process leaves a rich

researcharxiv-cs-cl
4 Jun 2026
Research

ReSGA: A Large Tail Risk Model for Learning Value-at-Risk and Expected Shortfall

DGX agent

arXiv:2606.04576v1 Announce Type: cross Abstract: Learning Value-at-Risk (VaR) and Expected Shortfall (ES) is important for managing financial risks effectively. Existing approaches with limited param

researcharxiv-cs-lg
4 Jun 2026
Research

SAID: Accelerating Diffusion-Based Language Models via Scaffold-Aware Iterative Decoding

DGX agent

arXiv:2606.04974v1 Announce Type: new Abstract: Diffusion large language models (DLLMs) enable non-autoregressive generation by iteratively denoising corrupted token sequences with bidirectional conte

researcharxiv-cs-cl
4 Jun 2026
Hardware

Scaling Novel Graph Generation via Lightweight Structure-Guided Autoregressive Models

DGX agent

arXiv:2606.04287v1 Announce Type: cross Abstract: Generating realistic and diverse graphs is a key problem in machine learning, with applications in molecular discovery, circuit design, cybersecurity,

hardwarearxiv-cs-ai
4 Jun 2026
Applications

The Mechanistic Emergence of Symbol Grounding in Language Models

DGX agent

arXiv:2510.13796v3 Announce Type: replace Abstract: Symbol grounding (Harnad, 1990) describes how symbols such as words acquire their meanings by connecting to real-world sensorimotor experiences. Rec

applicationsarxiv-cs-cl
4 Jun 2026
Research

T^star: Progressive Block Scaling for Masked Diffusion Language Models Through Trajectory Aware Reinforcement Learning

DGX agent

arXiv:2601.11214v5 Announce Type: replace Abstract: We present T^star, a simple TraceRL-based training curriculum for progressive block-size scaling in masked diffusion language models (MDMs). Startin

researcharxiv-cs-cl
4 Jun 2026
Research

An Attention-Based Denoising Model for Diffusion Weighted Imaging

DGX agent

arXiv:2606.03903v1 Announce Type: new Abstract: Diffusion-weighted imaging (DWI) is used for whole-body cancer screening, but it typically requires a long acquisition time. When the scan time is reduc

researcharxiv-cs-cv
3 Jun 2026
Research

Are Common Substructures Transferable? Riemannian Graph Foundation Model with Neural Vector Bundles

DGX agent

arXiv:2606.03270v1 Announce Type: cross Abstract: Foundation models have sparked a revolution via a pretraining-adaptation paradigm, with recent efforts extending this success to graphs. Unlike other

researcharxiv-cs-ai
3 Jun 2026
Tutorials

Causal Evidence of Stack Representations in Modeling Counter Languages Using Transformers

DGX agent

arXiv:2606.03398v1 Announce Type: cross Abstract: Formal languages have proven to be effective conduits to understand the inner mechanisms of transformers. Past work has shown that transformers traine

tutorialsarxiv-cs-ai
3 Jun 2026
Research

CTR-Sink: Attention Sink for Language Models in Click-Through Rate Prediction

DGX agent

arXiv:2508.03668v2 Announce Type: replace Abstract: Click-Through Rate (CTR) prediction, a core task in recommendation systems, estimates user click likelihood using historical behavioral data. Modeli

researcharxiv-cs-cl
3 Jun 2026
Local Ai

Distill-then-Replace: Efficient Task-Specific Hybrid Attention Model Construction

DGX agent

arXiv:2601.11667v2 Announce Type: replace-cross Abstract: Transformer architectures deliver state-of-the-art accuracy via dense full-attention, but their quadratic time and memory complexity with resp

local-aiarxiv-cs-ai
3 Jun 2026
Safety

GeoAlign: Beyond Semantics with State-Guided Spatial Alignment in VLA Models

DGX agent

arXiv:2606.03240v1 Announce Type: new Abstract: Current Vision--Language--Action (VLA) models often optimize for semantic grounding, whereas executable manipulation requires geometry-aware spatial ali

safetyarxiv-cs-ro
3 Jun 2026
Hardware

GRZO: Group-Relative Zeroth-Order Optimization for Large Language Model Fine-Tuning

DGX agent

arXiv:2606.02857v1 Announce Type: cross Abstract: Zeroth-order (ZO) optimization is a memory-efficient alternative to backpropagation for fine-tuning large language models, but its deployment is limit

hardwarearxiv-cs-ai
3 Jun 2026
Safety

Learning Unmasking Policies for Diffusion Language Models

DGX agent

arXiv:2512.09106v4 Announce Type: replace Abstract: Diffusion (Large) Language Models (dLLMs) now match the downstream performance of their autoregressive counterparts on many tasks, while holding the

safetyarxiv-cs-lg
3 Jun 2026
Tutorials

Oscillatory State-Space Models as Inductive Biases for Physics-Informed Neural PDE Solvers

DGX agent

arXiv:2606.02623v1 Announce Type: cross Abstract: Solving time-dependent partial differential equations (PDEs) is an important problem in computational science and engineering. Physics-informed neural

tutorialsarxiv-cs-ai
3 Jun 2026
Safety

Physical Plausibility Reasoning via HCM-GRPO: Empowering Compact Model for Superior Performance

DGX agent

arXiv:2511.10055v2 Announce Type: replace Abstract: The performance of image generation has been significantly improved in recent years. However, the study of image screening is rare, and its performa

safetyarxiv-cs-cv
3 Jun 2026
Local Ai

ReciNet: Reciprocal Space-Aware Long-Range Modeling for Crystalline Property Prediction

DGX agent

arXiv:2502.02748v4 Announce Type: replace Abstract: Predicting properties of crystals from their structures is a fundamental yet challenging task in materials science. Unlike molecules, crystal struct

local-aiarxiv-cs-lg
3 Jun 2026
Agents

Reinforcement Learning from Cross-domain Videos with Video Prediction Model

DGX agent

arXiv:2606.03201v1 Announce Type: cross Abstract: Reinforcement learning from expert videos across visually distinct domains is challenging due to the absence of reward signals and the presence of dom

agentsarxiv-cs-ai
3 Jun 2026
Research

Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling

DGX agent

arXiv:2606.03102v1 Announce Type: new Abstract: Test-time scaling improves the reasoning performance of large language models but incurs substantial cost in both total computation and latency. Existin

researcharxiv-cs-cl
3 Jun 2026
Hardware

Speedrunning Tabular Foundation Model Pretraining

DGX agent

arXiv:2606.03681v1 Announce Type: new Abstract: Pretraining cost is a major bottleneck for research on tabular foundation models, slowing the iteration cycle for new architectures, priors, and optimiz

hardwarearxiv-cs-lg
3 Jun 2026
Agents

Traj-Evolve: A Self-Evolving Multi-Agent System for Patient Trajectory Modeling in Lung Cancer Early Detection

DGX agent

arXiv:2606.02812v1 Announce Type: new Abstract: Modeling patient trajectories from longitudinal electronic health records (EHRs) requires reasoning over sparse, noisy, and long-context multimodal sequ

agentsarxiv-cs-ai
3 Jun 2026
Research

Typhoon: Towards an Effective Task-Specific Masking Strategy for Pre-trained Language Models

DGX agent

arXiv:2303.15619v2 Announce Type: replace-cross Abstract: The choice of which tokens to mask is a central, under-examined design decision in masked language modeling (MLM). Standard pretraining masks

researcharxiv-cs-ai
3 Jun 2026
Safety

An Enigma of Artificial Reason: Investigating the Production-Evaluation Gap in Large Reasoning Models

DGX agent

arXiv:2606.01462v1 Announce Type: new Abstract: Studies of human reasoning have shown that people are typically stronger at evaluating reasoning than producing it from scratch. In contrast, large reas

safetyarxiv-cs-ai
2 Jun 2026
Research

ART: Attention Run-time Termination for Efficient Large Language Model Decoding

DGX agent

arXiv:2606.00024v1 Announce Type: new Abstract: Long-context decoding in Large Language Models (LLMs) is severely constrained by the memory bandwidth required to fetch the extensive Key-Value (KV) cac

researcharxiv-cs-cl
2 Jun 2026
Safety

Behavior-Invariant Task Representation Learning with Transformer-based World Models for Offline Meta-Reinforcement Learning

DGX agent

arXiv:2606.00780v1 Announce Type: cross Abstract: Offline meta-reinforcement learning leverages static datasets to enable agents to generalize to unseen environments by combining offline efficiency wi

safetyarxiv-cs-ai
2 Jun 2026
← Previous
1…212213214215216…1038
Next →