AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

Adaptive Memory Momentum via a Model-Based Framework for Deep Learning Optimization

DGX agent

arXiv:2510.04988v3 Announce Type: replace Abstract: The vast majority of modern deep learning models are trained with momentum-based first-order optimizers. The momentum term governs the optimizer's m

researcharxiv-cs-lg
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Agentic Performance at the Edge: Insights from Benchmarking

DGX agent

arXiv:2605.10384v1 Announce Type: new Abstract: Agentic artificial intelligence (AI) is a natural fit for Internet of Things (IoT) and edge systems, but edge deployments are often constrained to model

model-releasesarxiv-cs-ai
12 May 2026
Research

Beta Sampling is All You Need: Efficient Image Generation Strategy for Diffusion Models using Stepwise Spectral Analysis

DGX agent

arXiv:2407.12173v2 Announce Type: replace-cross Abstract: Generative diffusion models have emerged as a powerful tool for high-quality image synthesis, yet their iterative nature demands significant c

researcharxiv-cs-ai
12 May 2026
Research

BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models

DGX agent

arXiv:2605.09134v1 Announce Type: new Abstract: Reinforcement learning for program repair is hindered by sparse execution feedback and coarse sequence-level rewards that obscure which edits actually f

researcharxiv-cs-ai
12 May 2026
Local Ai

C2L-Net: A Data-Driven Model for State-of-Charge Estimation of Lithium-Ion Batteries During Discharge

DGX agent

arXiv:2605.08653v1 Announce Type: new Abstract: Accurate state-of-charge (SOC) estimation is critical for the safe and efficient operation of lithium-ion batteries in battery management systems (BMS).

local-aiarxiv-cs-ai
12 May 2026
Research

CollabVR: Collaborative Video Reasoning with Vision-Language and Video Generation Models

DGX agent

arXiv:2605.08735v1 Announce Type: new Abstract: Recent 'Thinking with Video' approaches use Video Generation Models (VGMs) for visual reasoning by producing temporally coherent Chain-of-Frames as reas

researcharxiv-cs-cv
12 May 2026
Research

Contextual Plackett-Luce: An Efficient Neural Model for Probabilistic Sequence Selection under Ambiguity

DGX agent

arXiv:2605.09112v1 Announce Type: cross Abstract: Selecting a coherent sequence or subset of elements is a fundamental problem in structured prediction, arising in tasks such as detection, trajectory

researcharxiv-cs-ai
12 May 2026
Model Releases

Cross-Family Universality of Behavioral Axes via Anchor-Projected Representations

DGX agent

arXiv:2605.09875v1 Announce Type: new Abstract: Large language models from different families use different hidden dimensions, tokenizers, and training procedures, making behavioral directions difficu

model-releasesarxiv-cs-ai
12 May 2026
Safety

DAPE: Dynamic Non-uniform Alignment and Progressive Detail Enhancement Techniques for Improving the Performance of Efficient Visual Language Models

DGX agent

arXiv:2605.08902v1 Announce Type: cross Abstract: In recent years, pre-trained visual-linguistic models have demonstrated tremendous potential, becoming a crucial foundational framework for numerous d

safetyarxiv-cs-ai
12 May 2026
Research

Elucidating Representation Degradation Problem in Diffusion Model Training

DGX agent

arXiv:2605.10790v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success, yet their training remains inefficient due to a severe optimization bottleneck, which we term Represe

researcharxiv-cs-lg
12 May 2026
Research

Energy-based models for diagnostic reconstruction and analysis in a laboratory plasma device

DGX agent

arXiv:2605.08645v1 Announce Type: cross Abstract: Energy-based models (EBMs) provide a powerful and flexible way of learning a joint probability distribution over data by constructing an energy surfac

researcharxiv-cs-lg
12 May 2026
Agents

Enhancing Consistency Models for Multi-Agent Trajectory Prediction

DGX agent

arXiv:2605.08572v1 Announce Type: new Abstract: Diffusion models for multi-agent trajectory prediction are limited by iterative denoising, which causes inference latency that hinders their use in time

agentsarxiv-cs-cv
12 May 2026
Safety

Exploration-Driven Optimization for Test-Time Large Language Model Reasoning

DGX agent

arXiv:2605.09853v1 Announce Type: new Abstract: Post-training techniques combined with inference-time scaling significantly enhance the reasoning and alignment capabilities of large language models (L

safetyarxiv-cs-lg
12 May 2026
Research

ExtraVAR: Stage-Aware RoPE Remapping for Resolution Extrapolation in Visual Autoregressive Models

DGX agent

arXiv:2605.10045v1 Announce Type: new Abstract: Visual Autoregressive (VAR) models have emerged as a strong alternative to diffusion for image synthesis, yet their fixed training resolution prevents d

researcharxiv-cs-cv
12 May 2026
Model Releases

Flame3D: Zero-shot Compositional Reasoning of 3D Scenes with Agentic Language Models

DGX agent

arXiv:2605.09218v1 Announce Type: cross Abstract: 3D scene understanding spans reasoning about free space, object grounding, hypothetical object insertions, complex geometric relationships, and integr

model-releasesarxiv-cs-ai
12 May 2026
Agents

From Spark to Fire: Modeling and Mitigating Error Cascades in LLM-Based Multi-Agent Collaboration

DGX agent

arXiv:2603.04474v2 Announce Type: replace-cross Abstract: Large Language Model-based Multi-Agent Systems (LLM-MAS) are increasingly applied to complex collaborative scenarios. However, their collabora

agentsarxiv-cs-ai
12 May 2026
Research

Gate-and-Merge: Zero-shot Compositional Personalization of Vision Language Models

DGX agent

arXiv:2605.08702v1 Announce Type: cross Abstract: This paper tackles compositional personalization of vision-language models (VLMs). In this problem, multiple user-defined concepts must be recognized

researcharxiv-cs-ai
12 May 2026
Model Releases

Geometry-Aware Discretization Error of Diffusion Models

DGX agent

arXiv:2605.08392v1 Announce Type: new Abstract: Practical diffusion sampling is a numerical approximation problem: under a fixed inference budget, one must simulate a reverse-time ODE or SDE using onl

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Geospatial-Temporal Sensemaking of Remote Sensing Activity Detections with Multimodal Large Language Model

DGX agent

arXiv:2605.10739v1 Announce Type: cross Abstract: We introduce SMART-HC-VQA, a Sentinel-2-based visual question answering dataset derived from the IARPA SMART Heavy Construction dataset, designed for

model-releasesarxiv-cs-ai
12 May 2026
Safety

Large Language Models for Sequential Decision-Making: Improving In-Context Learning via Supervised Fine-Tuning

DGX agent

arXiv:2605.09009v1 Announce Type: cross Abstract: Large language models (LLMs) have shown remarkable in-context learning (ICL) capabilities, yet their potential for sequential decision-making remains

safetyarxiv-cs-ai
12 May 2026
Model Releases

Learning Less Is More: Premature Upper-Layer Attention Specialization Hurts Language Model Pretraining

DGX agent

arXiv:2605.10504v1 Announce Type: new Abstract: A causal-decoder block is hierarchical: lower layers build the residual basis that upper layers attend over. We identify a failure mode in GPT pretraini

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

LLM Agents Already Know When to Call Tools -- Even Without Reasoning

DGX agent

arXiv:2605.09252v1 Announce Type: new Abstract: Tool-augmented LLM agents tend to call tools indiscriminately, even when the model can answer directly. Each unnecessary call wastes API fees and latenc

model-releasesarxiv-cs-cl
12 May 2026
Safety

Metropolis-Adjusted Diffusion Models

DGX agent

arXiv:2605.09654v1 Announce Type: cross Abstract: Sampling from score-based diffusion models incurs bias due to both time discretisation and the approximation of the score function. A common strategy

safetyarxiv-cs-lg
12 May 2026
Safety

Mismatch-Aware Adaptive Constraint Tightening for Bicycle-Model Trajectory Optimization

DGX agent

arXiv:2605.09376v1 Announce Type: new Abstract: Trajectory optimization for autonomous vehicles usually relies on the kinematic bicycle model because of its computational simplicity. However, when the

safetyarxiv-cs-ro
12 May 2026
Research

Restoring Exploration after Post-Training: Latent Exploration Decoding for Large Reasoning Models

DGX agent

arXiv:2602.01698v3 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have recently achieved strong mathematical and code reasoning performance through Reinforcement Learning (RL) post-tra

researcharxiv-cs-cl
12 May 2026
Research

Rethinking Event-Based Object Dtection through Representation-Level Temporal Aggregation and Model-Level Hypergraph Reasoning

DGX agent

arXiv:2605.08825v1 Announce Type: new Abstract: Event cameras provide microsecond-level temporal resolution, low latency, and high dynamic range, offering potential for perception under fast motion an

researcharxiv-cs-cv
12 May 2026
Research

Self-Captioning Multimodal Interaction Tuning: Amplifying Exploitable Redundancies for Robust Vision Language Models

DGX agent

arXiv:2605.08145v1 Announce Type: cross Abstract: Current vision language models face hallucination and robustness issues against ambiguous or corrupted modalities. We hypothesize that these issues ca

researcharxiv-cs-ai
12 May 2026
Research

TARO: Temporal Adversarial Rectification Optimization Using Diffusion Models as Purifiers

DGX agent

arXiv:2605.08440v1 Announce Type: cross Abstract: Adversarial purification with diffusion models seeks to project adversarial examples back toward the data manifold, but balancing semantic preservatio

researcharxiv-cs-cv
12 May 2026
Model Releases

The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play

DGX agent

arXiv:2605.08427v1 Announce Type: new Abstract: Self-play red team is an established approach to improving AI safety in which different instances of the same model play attacker and defender roles in

model-releasesarxiv-cs-ai
12 May 2026
Agents

Towards Robust Surgical Automation via Digital Twin Representations from Foundation Models

DGX agent

arXiv:2409.13107v3 Announce Type: replace Abstract: Large language model-based (LLM) agents are emerging as a powerful enabler of robust embodied intelligence due to their capability of planning compl

agentsarxiv-cs-ro
12 May 2026
Safety

Training-Free Cultural Alignment of Large Language Models via Persona Disagreement

DGX agent

arXiv:2605.10843v1 Announce Type: cross Abstract: Large language models increasingly mediate decisions that turn on moral judgement, yet a growing body of evidence shows that their implicit preference

safetyarxiv-cs-ai
12 May 2026
Research

UM-Text: A Unified Multimodal Model for Image Understanding and Visual Text Editing

DGX agent

arXiv:2601.08321v3 Announce Type: replace Abstract: With the rapid advancement of image generation, visual text editing using natural language instructions has received increasing attention. The main

researcharxiv-cs-cv
12 May 2026
Applications

Unlocking air traffic flow prediction through microscopic aircraft-state modeling

DGX agent

arXiv:2605.10083v1 Announce Type: new Abstract: Short-term air traffic flow prediction in terminal airspace is essential for proactive air traffic management. Existing approaches predominantly model t

applicationsarxiv-cs-lg
12 May 2026
Research

UxSID: Semantic-Aware User Interests Modeling for Ultra-Long Sequence

DGX agent

arXiv:2605.09040v1 Announce Type: new Abstract: Modeling ultra-long user sequences involves a difficult trade-off between efficiency and effectiveness. While current paradigms rely on either item-spec

researcharxiv-cs-ai
12 May 2026
Agents

ViSRA: A Video-based Spatial Reasoning Agent for Multi-modal Large Language Models

DGX agent

arXiv:2605.10106v1 Announce Type: cross Abstract: Recent advances in Multi-modal Large Language Models (MLLMs) target 3D spatial intelligence, yet the progress has been largely driven by post-training

agentsarxiv-cs-ai
12 May 2026
Research

ViSurf: Visual Supervised-and-Reinforcement Fine-Tuning for Large Vision-and-Language Models

DGX agent

arXiv:2510.10606v4 Announce Type: replace Abstract: Post-training Large Vision-and-Language Models (LVLMs) typically involves Supervised Fine-Tuning (SFT) for knowledge injection or Reinforcement Lear

researcharxiv-cs-cv
12 May 2026
Agents

When Child Inherits: Modeling and Exploiting Subagent Spawn in Multi-Agent Networks

DGX agent

arXiv:2605.08460v1 Announce Type: cross Abstract: Since the official release of ChatGPT in 2022, large language models (LLMs) have rapidly evolved from chatbot-style interfaces into agentic systems th

agentsarxiv-cs-ai
12 May 2026
Safety

When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models

DGX agent

arXiv:2605.08245v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) increasingly power high-stakes applications, from medical imaging to autonomous systems, yet they routinely hallucinate,

safetyarxiv-cs-ai
12 May 2026
Safety

When More Parameters Hurt: Foundation Model Priors Amplify Worst-Client Disparity Under Extreme Federated Heterogeneity

DGX agent

arXiv:2605.08992v1 Announce Type: new Abstract: Federated learning (FL) is increasingly used to fine-tune foundation models (FMs) on distributed private data. The community largely assumes that large-

safetyarxiv-cs-lg
12 May 2026
Safety

Beyond Pairs: Your Language Model is Secretly Optimizing a Preference Graph

DGX agent

arXiv:2605.08037v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) aligns language models using pairwise preference comparisons, offering a simple and effective alternative to Rein

safetyarxiv-cs-ai
11 May 2026
Agents

CASCADE: Case-Based Continual Adaptation for Large Language Models During Deployment

DGX agent

arXiv:2605.06702v1 Announce Type: new Abstract: Large language models (LLMs) have become a central foundation of modern artificial intelligence, yet their lifecycle remains constrained by a rigid sepa

agentsarxiv-cs-ai
11 May 2026
Research

Code Generation and Conic Constraints for Model-Predictive Control on Microcontrollers with Conic-TinyMPC

DGX agent

arXiv:2403.18149v3 Announce Type: replace Abstract: Model-predictive control (MPC) is a state-of-the-art control method for constrained robotic systems, yet deployment on resource-limited hardware rem

researcharxiv-cs-ro
11 May 2026
Safety

Cognitive Agent Compilation for Explicit Problem Solver Modeling

DGX agent

arXiv:2605.07040v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for tutoring, feedback generation, and content creation, but their broad pretraining makes them hard to c

safetyarxiv-cs-ai
11 May 2026
Research

Conservative Flows: A New Paradigm of Generative Models

DGX agent

arXiv:2605.06905v1 Announce Type: new Abstract: Modern generative modeling is dominated by transport from a noise prior to data. We propose an alternative paradigm in which generation is performed by

researcharxiv-cs-lg
11 May 2026
Model Releases

Dataset Watermarking for Closed LLMs with Provable Detection

DGX agent

arXiv:2605.06865v1 Announce Type: new Abstract: Large language models (LLMs) are pre-trained and post-trained on vast amounts of loosely curated data, raising the possibility that these models may hav

model-releasesarxiv-cs-lg
11 May 2026
Applications

Emergence of Distortions in High-Dimensional Guided Diffusion Models

DGX agent

arXiv:2602.00716v4 Announce Type: replace-cross Abstract: Classifier-free guidance (CFG) is the de facto standard for conditional sampling in diffusion models, yet it often reduces sample diversity. U

applicationsarxiv-cs-lg
11 May 2026
Research

Equivalence of Coarse and Fine-Grained Models for Learning with Distribution Shift

DGX agent

arXiv:2605.07005v1 Announce Type: cross Abstract: Recent work on provably efficient algorithms for learning with distribution shift has focused on two models: PQ learning (Goldwasser et al. (2020)) an

researcharxiv-cs-lg
11 May 2026
Local Ai

FLAM: Evaluating Model Performance with Aggregatable Measures in Federated Learning

DGX agent

arXiv:2605.07962v1 Announce Type: new Abstract: Performance evaluation is essential for assessing the quality of machine learning (ML) models and guiding deployment decisions. In federated learning (F

local-aiarxiv-cs-lg
11 May 2026
← Previous
1…189190191192193…1030
Next →