AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

Doubly Outlier-Robust Online Infinite Hidden Markov Model

DGX agent

arXiv:2604.14322v1 Announce Type: cross Abstract: We derive a robust update rule for the online infinite hidden Markov model (iHMM) for when the streaming data contains outliers and the model is missp

researcharxiv-cs-lg
17 Apr 2026
Applications
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

HARNESS: Lightweight Distilled Arabic Speech Foundation Models

DGX agent

arXiv:2604.14186v1 Announce Type: cross Abstract: Large self-supervised speech (SSL) models achieve strong downstream performance, but their size limits deployment in resource-constrained settings. We

applicationsarxiv-cs-cl
17 Apr 2026
Tutorials

How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data

DGX agent

arXiv:2604.14164v1 Announce Type: new Abstract: A widely adopted strategy for model enhancement is to use synthetic data generated by a stronger model for supervised fine-tuning (SFT). However, for em

tutorialsarxiv-cs-cl
17 Apr 2026
Local Ai

Large Vision Model-Guided Masked Low-Rank Approximation for Ground-Roll Attenuation

DGX agent

arXiv:2604.00998v2 Announce Type: replace Abstract: Ground roll is a common type of coherent noise in seismic records, and its attenuation remains challenging due to its substantial overlap with usefu

local-aiarxiv-cs-cv
17 Apr 2026
Safety

LeapAlign: Post-Training Flow Matching Models at Any Generation Step by Building Two-Step Trajectories

DGX agent

arXiv:2604.15311v1 Announce Type: new Abstract: This paper focuses on the alignment of flow matching models with human preferences. A promising way is fine-tuning by directly backpropagating reward gr

safetyarxiv-cs-cv
17 Apr 2026
Agents

MCPThreatHive: Automated Threat Intelligence for Model Context Protocol Ecosystems

DGX agent

arXiv:2604.13849v1 Announce Type: cross Abstract: The rapid proliferation of Model Context Protocol (MCP)-based agentic systems has introduced a new category of security threats that existing framewor

agentsarxiv-cs-ai
17 Apr 2026
Research

SPaCe: Unlocking Sample-Efficient Large Language Models Training With Self-Pace Curriculum Learning

DGX agent

arXiv:2508.05015v2 Announce Type: replace Abstract: Large language models (LLMs) have shown strong reasoning capabilities when fine-tuned with reinforcement learning (RL). However, such methods requir

researcharxiv-cs-lg
17 Apr 2026
Model Releases

TennisTV: Do Multimodal Large Language Models Understand Tennis Rallies?

DGX agent

arXiv:2509.15602v5 Announce Type: replace Abstract: Multimodal large language models (MLLMs) excel at general video understanding but struggle with fast, high-frequency sports like tennis, where rally

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

The Mirror Design Pattern: Strict Data Geometry over Model Scale for Prompt Injection Detection

DGX agent

arXiv:2603.11875v2 Announce Type: replace-cross Abstract: Prompt injection defenses are often framed as semantic understanding problems and delegated to increasingly large neural detectors. For the fi

model-releasesarxiv-cs-ai
17 Apr 2026
Research

VisPCO: Visual Token Pruning Configuration Optimization via Budget-Aware Pareto-Frontier Learning for Vision-Language Models

DGX agent

arXiv:2604.15188v1 Announce Type: new Abstract: Visual token pruning methods effectively mitigate the quadratic computational growth caused by processing high-resolution images and video frames in vis

researcharxiv-cs-cv
17 Apr 2026
Local Ai

A KL Lens on Quantization: Fast, Forward-Only Sensitivity for Mixed-Precision SSM-Transformer Models

DGX agent

arXiv:2604.13440v1 Announce Type: new Abstract: Deploying Large Language Models (LLMs) on edge devices faces severe computational and memory constraints, limiting real-time processing and on-device in

local-aiarxiv-cs-lg
16 Apr 2026
Research

A1: A Fully Transparent Open-Source, Adaptive and Efficient Truncated Vision-Language-Action Model

DGX agent

arXiv:2604.05672v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have emerged as a powerful paradigm for open-world robot manipulation, but their practical deployment is often c

researcharxiv-cs-ro
16 Apr 2026
Applications

Abstract 3D Perception for Spatial Intelligence in Vision-Language Models

DGX agent

arXiv:2511.10946v3 Announce Type: replace Abstract: Vision-language models (VLMs) struggle with 3D-related tasks such as spatial cognition and physical understanding, which are crucial for real-world

applicationsarxiv-cs-cv
16 Apr 2026
Research

Adaptive Learning via Off-Model Training and Importance Sampling for Fully Non-Markovian Optimal Stochastic Control. Complete version

DGX agent

arXiv:2604.13147v1 Announce Type: cross Abstract: This paper studies continuous-time stochastic control problems whose controlled states are fully non-Markovian and depend on unknown model parameters.

researcharxiv-cs-lg
16 Apr 2026
Tutorials

Addressing Overthinking in Large Vision-Language Models via Gated Perception-Reasoning Optimization

DGX agent

arXiv:2601.04442v2 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) have exhibited strong reasoning capabilities through chain-of-thought mechanisms that generate step-by-st

tutorialsarxiv-cs-cl
16 Apr 2026
Applications

Before the First Token: Scale-Dependent Emergence of Hallucination Signals in Autoregressive Language Models

DGX agent

arXiv:2604.13068v1 Announce Type: new Abstract: When do large language models decide to hallucinate? Despite serious consequences in healthcare, law, and finance, few formal answers exist. Recent work

applicationsarxiv-cs-cl
16 Apr 2026
Research

Better and Worse with Scale: How Contextual Entrainment Diverges with Model Size

DGX agent

arXiv:2604.13275v1 Announce Type: new Abstract: Larger language models become simultaneously better and worse at handling contextual information -- better at ignoring false claims, worse at ignoring i

researcharxiv-cs-cl
16 Apr 2026
Model Releases

Decoding the Delta: Unifying Remote Sensing Change Detection and Understanding with Multimodal Large Language Models

DGX agent

arXiv:2604.14044v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) excel in general vision-language tasks, their application to remote sensing change understanding is hinde

model-releasesarxiv-cs-cv
16 Apr 2026
Research

Democratising Pathology Co-Pilots: An Open Pipeline and Dataset for Whole-Slide Vision-Language Modelling

DGX agent

arXiv:2512.17326v2 Announce Type: replace Abstract: Vision-language models (VLMs) have the potential to become co-pilots for pathologists. However, most VLMs either focus on small regions of interest

researcharxiv-cs-cv
16 Apr 2026
Applications

Feed-Forward 3D Scene Modeling: A Problem-Driven Perspective

DGX agent

arXiv:2604.14025v1 Announce Type: new Abstract: Reconstructing 3D representations from 2D inputs is a fundamental task in computer vision and graphics, serving as a cornerstone for understanding and i

applicationsarxiv-cs-cv
16 Apr 2026
Safety

Heavy-Tailed Class-Conditional Priors for Long-Tailed Generative Modeling

DGX agent

arXiv:2509.02154v2 Announce Type: replace-cross Abstract: Variational Autoencoders (VAEs) with global priors trained under an imbalanced empirical class distribution can lead to underrepresentation of

safetyarxiv-cs-cv
16 Apr 2026
Model Releases

LaoBench: A Large-Scale Multidimensional Lao Benchmark for Large Language Models

DGX agent

arXiv:2511.11334v3 Announce Type: replace Abstract: The rapid advancement of large language models (LLMs) has not been matched by their evaluation in low-resource languages, especially Southeast Asian

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Mitigating Barren Plateaus in Quantum Denoising Diffusion Probabilistic Model

DGX agent

arXiv:2512.06695v2 Announce Type: replace Abstract: Quantum generative models exploit quantum superposition and entanglement to enhance learning efficiency for both classical and quantum data. Recentl

researcharxiv-cs-lg
16 Apr 2026
Model Releases

MOONSHOT : A Framework for Multi-Objective Pruning of Vision and Large Language Models

DGX agent

arXiv:2604.13287v1 Announce Type: new Abstract: Weight pruning is a common technique for compressing large neural networks. We focus on the challenging post-training one-shot setting, where a pre-trai

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

MulDimIF: A Multi-Dimensional Constraint Framework for Evaluating and Improving Instruction Following in Large Language Models

DGX agent

arXiv:2505.07591v2 Announce Type: replace Abstract: Instruction following refers to the ability of large language models (LLMs) to generate outputs that satisfy all specified constraints. Existing res

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Reward Design for Physical Reasoning in Vision-Language Models

DGX agent

arXiv:2604.13993v1 Announce Type: cross Abstract: Physical reasoning over visual inputs demands tight integration of visual perception, domain knowledge, and multi-step symbolic inference. Yet even st

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

UniBlendNet: Unified Global, Multi-Scale, and Region-Adaptive Modeling for Ambient Lighting Normalization

DGX agent

arXiv:2604.13383v1 Announce Type: new Abstract: Ambient Lighting Normalization (ALN) aims to restore images degraded by complex, spatially varying illumination conditions. Existing methods, such as IF

model-releasesarxiv-cs-cv
16 Apr 2026
Research

VLMs Need Words: Vision Language Models Ignore Visual Detail In Favor of Semantic Anchors

DGX agent

arXiv:2604.02486v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have achieved impressive performance across a wide range of multimodal tasks. However, they often fail on tasks

researcharxiv-cs-cl
16 Apr 2026
Model Releases

When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?

DGX agent

arXiv:2503.23137v2 Announce Type: replace-cross Abstract: Understanding humor-particularly when it involves complex, contradictory narratives that require comparative reasoning-remains a significant c

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

A General Model for Deepfake Speech Detection: Diverse Bonafide Resources or Diverse AI-Based Generators

DGX agent

arXiv:2603.27557v2 Announce Type: replace-cross Abstract: In this paper, we analyze two main factors of Bonafide Resource (BR) or AI-based Generator (AG) which affect the performance and the generalit

model-releasesarxiv-cs-ai
15 Apr 2026
Research

Conflated Inverse Modeling to Generate Diverse and Temperature-Change Inducing Urban Vegetation Patterns

DGX agent

arXiv:2604.13028v1 Announce Type: new Abstract: Urban areas are increasingly vulnerable to thermal extremes driven by rapid urbanization and climate change. Traditionally, thermal extremes have been m

researcharxiv-cs-cv
15 Apr 2026
Research

Dynamic Modeling and Robust Gait Optimization of a Compliant Worm Robot

DGX agent

arXiv:2604.12031v1 Announce Type: new Abstract: Worm-inspired robots provide an effective locomotion strategy for constrained environments by combining cyclic body deformation with alternating anchori

researcharxiv-cs-ro
15 Apr 2026
Model Releases

GroupKAN: Efficient Kolmogorov-Arnold Networks via Grouped Spline Modeling

DGX agent

arXiv:2511.05477v2 Announce Type: replace Abstract: Medical image segmentation demands models that achieve high accuracy while maintaining computational efficiency and clinical interpretability. While

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

HazardArena: Evaluating Semantic Safety in Vision-Language-Action Models

DGX agent

arXiv:2604.12447v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models inherit rich world knowledge from vision-language backbones and acquire executable skills via action demonstrations.

model-releasesarxiv-cs-ro
15 Apr 2026
Applications

LLM-HYPER: Generative CTR Modeling for Cold-Start Ad Personalization via LLM-Based Hypernetworks

DGX agent

arXiv:2604.12096v1 Announce Type: new Abstract: On online advertising platforms, newly introduced promotional ads face the cold-start problem, as they lack sufficient user feedback for model training.

applicationsarxiv-cs-ai
15 Apr 2026
Safety

MODIX: A Training-Free Multimodal Information-Driven Positional Index Scaling for Vision-Language Models

DGX agent

arXiv:2604.12537v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved remarkable progress in multimodal understanding, yet their positional encoding mechanisms remain suboptima

safetyarxiv-cs-ai
15 Apr 2026
Model Releases

Narrative over Numbers: The Identifiable Victim Effect and its Amplification Under Alignment and Reasoning in Large Language Models

DGX agent

arXiv:2604.12076v1 Announce Type: cross Abstract: The Identifiable Victim Effect (IVE) - the tendency to allocate greater resources to a specific, narratively described victim than to a statistically

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: Professional Image Quality Assessment (Track 1)

DGX agent

arXiv:2604.12512v1 Announce Type: cross Abstract: In this paper, we present an overview of the NTIRE 2026 challenge on the 3rd Restore Any Image Model in the Wild, specifically focusing on Track 1: Pr

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models

DGX agent

arXiv:2604.12582v1 Announce Type: new Abstract: Recent Video Large Language Models (Video-LLMs) have demonstrated strong capability in video understanding, yet they still suffer from hallucinations. E

safetyarxiv-cs-cv
15 Apr 2026
Model Releases

Understanding or Memorizing? A Case Study of German Definite Articles in Language Models

DGX agent

arXiv:2601.09313v2 Announce Type: replace-cross Abstract: Language models perform well on grammatical agreement, but it is unclear whether this reflects rule-based generalization or memorization. We s

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Variation in Verification: Understanding Verification Dynamics in Large Language Models

DGX agent

arXiv:2509.17995v2 Announce Type: replace-cross Abstract: Recent advances have shown that scaling test-time computation enables large language models (LLMs) to solve increasingly complex problems acro

model-releasesarxiv-cs-ai
15 Apr 2026
Research

Visual Diffusion Models are Geometric Solvers

DGX agent

arXiv:2510.21697v2 Announce Type: replace Abstract: In this paper we show that visual diffusion models can serve as effective geometric solvers: they can directly reason about geometric problems by wo

researcharxiv-cs-cv
15 Apr 2026
Tutorials

Why Did Apple Fall: Evaluating Curiosity in Large Language Models

DGX agent

arXiv:2510.20635v2 Announce Type: replace-cross Abstract: Curiosity serves as a pivotal conduit for human beings to discover and learn new knowledge. Recent advancements of large language models (LLMs

tutorialsarxiv-cs-ai
15 Apr 2026
Model Releases

BadGraph: A Backdoor Attack Against Latent Diffusion Model for Text-Guided Graph Generation

DGX agent

arXiv:2510.20792v4 Announce Type: replace-cross Abstract: The rapid progress of graph generation has raised new security concerns, particularly regarding backdoor vulnerabilities. While prior work has

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Belief-Aware VLM Model for Human-like Reasoning

DGX agent

arXiv:2604.09686v1 Announce Type: new Abstract: Traditional neural network models for intent inference rely heavily on observable states and struggle to generalize across diverse tasks and dynamic env

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Can Large Language Models Infer Causal Relationships from Real-World Text?

DGX agent

arXiv:2505.18931v4 Announce Type: replace Abstract: Understanding and inferring causal relationships from texts is a core aspect of human cognition and is essential for advancing large language models

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models

DGX agent

arXiv:2503.21380v3 Announce Type: replace Abstract: The rapid advancement of large reasoning models has saturated existing math benchmarks, underscoring the urgent need for more challenging evaluation

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Computational Lesions in Multilingual Language Models Separate Shared and Language-specific Brain Alignment

DGX agent

arXiv:2604.10627v1 Announce Type: cross Abstract: How the brain supports language across different languages is a basic question in neuroscience and a useful test for multilingual artificial intellige

model-releasesarxiv-cs-ai
14 Apr 2026
← Previous
1…9495969798…1030
Next →