AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

PRAGMA: Revolut Foundation Model

DGX agent

arXiv:2604.08649v1 Announce Type: cross Abstract: Modern financial systems generate vast quantities of transactional and event-level data that encode rich economic signals. This paper presents PRAGMA,

researcharxiv-cs-cl
13 Apr 2026
Model Releases

QoS-QoE Translation with Large Language Model

DGX agent

arXiv:2604.08703v1 Announce Type: cross Abstract: QoS-QoE translation is a fundamental problem in multimedia systems because it characterizes how measurable system and network conditions affect user-p

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Robust Reasoning Benchmark

DGX agent

arXiv:2604.08571v1 Announce Type: cross Abstract: While Large Language Models (LLMs) achieve high performance on standard mathematical benchmarks, their underlying reasoning processes remain highly ov

model-releasesarxiv-cs-ai
13 Apr 2026
Local Ai

Tango: Taming Visual Signals for Efficient Video Large Language Models

DGX agent

arXiv:2604.09547v1 Announce Type: new Abstract: Token pruning has emerged as a mainstream approach for developing efficient Video Large Language Models (Video LLMs). This work revisits and advances th

local-aiarxiv-cs-cv
13 Apr 2026
Research

Token Reduction via Local and Global Contexts Optimization for Efficient Video Large Language Models

DGX agent

arXiv:2603.01400v2 Announce Type: replace Abstract: Video Large Language Models (VLLMs) demonstrate strong video understanding but suffer from inefficiency due to redundant visual tokens. Existing pru

researcharxiv-cs-cv
13 Apr 2026
Tutorials

What Matters in Virtual Try-Off? Dual-UNet Diffusion Model For Garment Reconstruction

DGX agent

arXiv:2604.08716v1 Announce Type: new Abstract: Virtual Try-On (VTON) has seen rapid advancements, providing a strong foundation for generative fashion tasks. However, the inverse problem, Virtual Try

tutorialsarxiv-cs-cv
13 Apr 2026
Safety

A systematic framework for generating novel experimental hypotheses from language models

DGX agent

arXiv:2408.05086v3 Announce Type: replace Abstract: Neural language models (LMs) have been shown to capture complex linguistic patterns, yet their utility in understanding human language and more broa

safetyarxiv-cs-cl
10 Apr 2026
Research

ABMAMBA: Multimodal Large Language Model with Aligned Hierarchical Bidirectional Scan for Efficient Video Captioning

DGX agent

arXiv:2604.08050v1 Announce Type: new Abstract: In this study, we focus on video captioning by fully open multimodal large language models (MLLMs). The comprehension of visual sequences is challenging

researcharxiv-cs-cv
10 Apr 2026
Research

AFL: A Single-Round Analytic Approach for Federated Learning with Pre-trained Models

DGX agent

arXiv:2405.16240v3 Announce Type: replace Abstract: In this paper, we introduce analytic federated learning (AFL), a new training paradigm that brings analytical (i.e., closed-form) solutions to the f

researcharxiv-cs-lg
10 Apr 2026
Tutorials

Attribution-Driven Explainable Intrusion Detection with Encoder-Based Large Language Models

DGX agent

arXiv:2604.06266v1 Announce Type: cross Abstract: Software-Defined Networking (SDN) improves network flexibility but also increases the need for reliable and interpretable intrusion detection. Large L

tutorialsarxiv-cs-ai
10 Apr 2026
Research

Beyond the Mean: Modelling Annotation Distributions in Continuous Affect Prediction

DGX agent

arXiv:2604.07198v1 Announce Type: new Abstract: Emotion annotation is inherently subjective and cognitively demanding, producing signals that reflect diverse perceptions across annotators rather than

researcharxiv-cs-lg
10 Apr 2026
Model Releases

DISSECT: Diagnosing Where Vision Ends and Language Priors Begin in Scientific VLMs

DGX agent

arXiv:2604.06250v1 Announce Type: cross Abstract: When asked to describe a molecular diagram, a Vision-Language Model correctly identifies ``a benzene ring with an -OH group.'' When asked to reason ab

model-releasesarxiv-cs-ai
10 Apr 2026
Research

'Don't Do That!': Guiding Embodied Systems through Large Language Model-based Constraint Generation

DGX agent

arXiv:2506.04500v3 Announce Type: replace-cross Abstract: Recent advancements in large language models (LLMs) have spurred interest in robotic navigation that incorporates complex spatial, mathematica

researcharxiv-cs-ro
10 Apr 2026
Research

Energy-Regularized Spatial Masking: A Novel Approach to Enhancing Robustness and Interpretability in Vision Models

DGX agent

arXiv:2604.06893v1 Announce Type: cross Abstract: Deep convolutional neural networks achieve remarkable performance by exhaustively processing dense spatial feature maps, yet this brute-force strategy

researcharxiv-cs-lg
10 Apr 2026
Model Releases

Flow Motion Policy: Manipulator Motion Planning with Flow Matching Models

DGX agent

arXiv:2604.07084v1 Announce Type: cross Abstract: Open-loop end-to-end neural motion planners have recently been proposed to improve motion planning for robotic manipulators. These methods enable plan

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Hallucination as output-boundary misclassification: a composite abstention architecture for language models

DGX agent

arXiv:2604.06195v1 Announce Type: cross Abstract: Large language models often produce unsupported claims. We frame this as a misclassification error at the output boundary, where internally generated

researcharxiv-cs-ai
10 Apr 2026
Research

Hierarchical Feature Learning for Medical Point Clouds via State Space Model

DGX agent

arXiv:2504.13015v3 Announce Type: replace Abstract: Deep learning-based point cloud modeling has been widely investigated as an indispensable component of general shape analysis. Recently, transformer

researcharxiv-cs-cv
10 Apr 2026
Research

Incentive-Aware Multi-Fidelity Optimization for Generative Advertising in Large Language Models

DGX agent

arXiv:2604.06263v1 Announce Type: cross Abstract: Generative advertising in large language model (LLM) responses requires optimizing sponsorship configurations under two strict constraints: the strate

researcharxiv-cs-ai
10 Apr 2026
Tutorials

Learning is Forgetting: LLM Training As Lossy Compression

DGX agent

arXiv:2604.07569v1 Announce Type: cross Abstract: Despite the increasing prevalence of large language models (LLMs), we still have a limited understanding of how their representational spaces are stru

tutorialsarxiv-cs-cl
10 Apr 2026
Safety

LINE: LLM-based Iterative Neuron Explanations for Vision Models

DGX agent

arXiv:2604.08039v1 Announce Type: new Abstract: Interpreting the concepts encoded by individual neurons in deep neural networks is a crucial step towards understanding their complex decision-making pr

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning

DGX agent

arXiv:2506.18841v3 Announce Type: replace-cross Abstract: Ultra-long generation by large language models (LLMs) is a widely demanded scenario, yet it remains a significant challenge due to their maxim

model-releasesarxiv-cs-ai
10 Apr 2026
Tutorials

ODE-free Neural Flow Matching for One-Step Generative Modeling

DGX agent

arXiv:2604.06413v1 Announce Type: new Abstract: Diffusion and flow matching models generate samples by learning time-dependent vector fields whose integration transports noise to data, requiring tens

tutorialsarxiv-cs-lg
10 Apr 2026
Agents

One Life to Learn: Inferring Symbolic World Models for Stochastic Environments from Unguided Exploration

DGX agent

arXiv:2510.12088v2 Announce Type: replace Abstract: Symbolic world modeling requires inferring and representing an environment's transitional dynamics as an executable program. Prior work has focused

agentsarxiv-cs-ai
10 Apr 2026
Model Releases

Rethinking Entropy Allocation in LLM-based ASR: Understanding the Dynamics between Speech Encoders and LLMs

DGX agent

arXiv:2604.08003v1 Announce Type: cross Abstract: Integrating large language models (LLMs) into automatic speech recognition (ASR) has become a dominant paradigm. Although recent LLM-based ASR models

model-releasesarxiv-cs-cl
10 Apr 2026
Research

STQuant: Spatio-Temporal Adaptive Framework for Optimizer Quantization in Large Multimodal Model Training

DGX agent

arXiv:2604.06836v2 Announce Type: new Abstract: Quantization is an effective way to reduce the memory cost of large-scale model training. However, most existing methods adopt fixed-precision policies,

researcharxiv-cs-lg
10 Apr 2026
Model Releases

Visual prompting reimagined: The power of the Activation Prompts

DGX agent

arXiv:2604.06440v1 Announce Type: cross Abstract: Visual prompting (VP) has emerged as a popular method to repurpose pretrained vision models for adaptation to downstream tasks. Unlike conventional mo

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

A Probe Direction Is a Property of Its Prompt

DGX agent

arXiv:2608.13329v1 Announce Type: new Abstract: A model that behaves differently when it senses it is being tested would undermine the evaluations we rely on, so recent work has sought to read that se

model-releasesarxiv-cs-lg
14 Aug 2026
Research

Capstan-driven Continuum Surgical Robot: Design, Modeling, and Perception

DGX agent

arXiv:2608.13396v1 Announce Type: new Abstract: Shape and force sensing have long been critical bottlenecks in the development of compact capstan-driven continuum surgical robots, primarily due to the

researcharxiv-cs-ro
14 Aug 2026
Research

Demand Transfer Estimation at Scale via Restricted Logit Modeling

DGX agent

arXiv:2608.12680v1 Announce Type: cross Abstract: Item demand forecasting is an integral component of store assortment optimization. Existing literature focuses on learning a suitable customer choice

researcharxiv-cs-ai
14 Aug 2026
Safety

DiffGRM: Diffusion-based Generative Recommendation Model

DGX agent

arXiv:2510.21805v2 Announce Type: replace-cross Abstract: Generative recommendation (GR) is an emerging paradigm that represents each item via a tokenizer as an n-digit semantic ID (SID) and predicts

safetyarxiv-cs-ai
14 Aug 2026
Research

From Observation to Intervention: Memory in Brains and Large Language Models

DGX agent

arXiv:2608.12377v1 Announce Type: cross Abstract: Brains and large language models (LLMs) are fundamentally different memory systems, but they can be compared through shared functional questions: wher

researcharxiv-cs-ai
14 Aug 2026
Safety

Intern-S2-Preview: Scientific Agentic Foundation Model

DGX agent

arXiv:2608.13505v1 Announce Type: cross Abstract: Scientific discovery increasingly requires AI systems that can reason over scientific evidence of heterogeneous modalities, interact with scientific t

safetyarxiv-cs-cl
14 Aug 2026
Safety

JailWAM: Jailbreaking World Action Models in Robot Control

DGX agent

arXiv:2604.05498v2 Announce Type: replace Abstract: World Action Models (WAMs) have emerged as a promising paradigm for robotic manipulation, enabling physical interaction across diverse tasks and env

safetyarxiv-cs-ro
14 Aug 2026
Research

Structure-aware Riemannian Growth Fields for 4D Plant Modeling

DGX agent

arXiv:2608.13007v1 Announce Type: new Abstract: In this paper, we introduce a novel framework for 4D plant growth modeling that reconstructs the continuous geometric and topological evolution of plant

researcharxiv-cs-cv
14 Aug 2026
Research

A Reality Check of Language Models as Formalizers on Constraint Satisfaction Problems

DGX agent

arXiv:2505.13252v5 Announce Type: replace Abstract: Recent work shows superior performance when using large language models (LLMs) as formalizers instead of as end-to-end solvers for symbolic reasonin

researcharxiv-cs-cl
13 Aug 2026
Research

Chemically Meaningful Textualization Enables Explainable Validation of Metal-Organic Frameworks by Large Language Models

DGX agent

arXiv:2608.11283v1 Announce Type: cross Abstract: Computation-ready metal-organic framework (MOF) databases are essential for high-throughput screening, yet many reported crystal structures remain che

researcharxiv-cs-ai
13 Aug 2026
Local Ai

Class Activation Mapping in Explainable Computer Vision: A Method-Centered Review of CNN, Transformer, and Foundation-Model-Era Visual Explanations

DGX agent

arXiv:2608.12299v1 Announce Type: cross Abstract: Class activation mapping (CAM) is one of the most widely used visual explanation families in explainable artificial intelligence. Its purpose is intui

local-aiarxiv-cs-ai
13 Aug 2026
Tutorials

Fine-Tuning Generative Models for Extreme Events via CVaR-Penalized Wasserstein Gradient Flows

DGX agent

arXiv:2608.11544v1 Announce Type: cross Abstract: We propose CVaR-penalized Generative Particle Algorithm (CVaR-GPA), a robust, tail-agnostic algorithm for fine-tuning generative models to learn heavy

tutorialsarxiv-cs-lg
13 Aug 2026
Model Releases

Governing Agentic AI in FinTech

DGX agent

arXiv:2608.11344v1 Announce Type: cross Abstract: Financial institutions are delegating consequential decisions to agentic AI systems that decompose goals, coordinate models and tools, and act with li

model-releasesarxiv-cs-ai
13 Aug 2026
Safety

Grounding Large Language Models as Generalizable Policies in Network Control

DGX agent

arXiv:2512.11839v2 Announce Type: replace Abstract: Designing generalizable control policies that operate reliably under changing conditions is essential for robust network services in modern digital

safetyarxiv-cs-lg
13 Aug 2026
Applications

Keep the Future, Drop the Rollout: RIFT for World Action Models

DGX agent

arXiv:2608.11521v1 Announce Type: cross Abstract: World action models (WAMs) condition robot actions on predicted futures, but iterative video rollout increases deployment latency. We ask whether acti

applicationsarxiv-cs-ai
13 Aug 2026
Model Releases

Latent variable models for simultaneous EOV identification and removal in population-based SHM

DGX agent

arXiv:2608.11995v1 Announce Type: cross Abstract: The robust treatment of environmental and operational variability (EOV) is an open challenge in population-based structural health monitoring (PBSHM).

model-releasesarxiv-cs-lg
13 Aug 2026
Model Releases

Self-Evolving Embodied Agents via Skill-Harness Evolution

DGX agent

arXiv:2608.11350v1 Announce Type: new Abstract: Embodied agents are increasingly built as systems around foundation models, where performance depends not only on model weights but also on the skills,

model-releasesarxiv-cs-cl
13 Aug 2026
Agents

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models

DGX agent

arXiv:2608.10278v1 Announce Type: new Abstract: Spatial understanding is fundamental to embodied intelligence, underpinning applications such as robotic manipulation, embodied navigation, and autonomo

agentsarxiv-cs-cv
12 Aug 2026
Applications

Covert Visual Prompt Injection against Commercial Multimodal Large Language Models

DGX agent

arXiv:2603.29418v2 Announce Type: replace-cross Abstract: Although multimodal large language models (MLLMs) are increasingly deployed in real-world applications, their instruction-following behavior l

applicationsarxiv-cs-ai
12 Aug 2026
Model Releases

Diffract: Spectral View of LLM Domain Adaptation

DGX agent

arXiv:2608.10850v1 Announce Type: new Abstract: We study continual pre-training (CPT) as a mechanism for adapting general-purpose large language models to specialized domains: mathematics, instruction

model-releasesarxiv-cs-lg
12 Aug 2026
Research

Dynamic Context Adapters: Efficiently Infusing History into Vision-and-Language Models

DGX agent

arXiv:2608.10525v1 Announce Type: cross Abstract: Historical context integration presents a fundamental challenge for Vision-Language Models (VLMs) in sequential decision-making tasks. Current VLMs pr

researcharxiv-cs-ai
12 Aug 2026
Model Releases

How Robust Are LLMs to Vietnamese Dialects?

DGX agent

arXiv:2608.10414v1 Announce Type: new Abstract: Large Language Models (LLMs) are typically evaluated on standard written Vietnamese, yet everyday communication frequently involves regional dialects th

model-releasesarxiv-cs-cl
12 Aug 2026
← Previous
1…169170171172173…1030
Next →