AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

SolarFCD: A Large-Scale Dataset and Benchmark for Solar Fault Classification in Photovoltaic Systems

DGX agent

arXiv:2604.23662v1 Announce Type: new Abstract: The increasing global deployment of solar photovoltaic (PV) systems needs robust, scalable, and automated inspection technologies capable of detecting a

model-releasesarxiv-cs-cv
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SPAGS: Sparse-View Articulated Object Reconstruction from Single State via Planar Gaussian Splatting

DGX agent

arXiv:2511.17092v4 Announce Type: replace Abstract: Articulated objects are ubiquitous in daily environments, and their 3D reconstruction holds great significance across various fields. However, exist

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Spatiotemporal Degradation-Aware 3D Gaussian Splatting for Realistic Underwater Scene Reconstruction

DGX agent

arXiv:2604.23551v1 Announce Type: new Abstract: Reconstructing realistic underwater scenes from underwater video remains a meaningful yet challenging task in the multimedia domain. The inherent spatio

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SpecRLBench: A Benchmark for Generalization in Specification-Guided Reinforcement Learning

DGX agent

arXiv:2604.24729v1 Announce Type: new Abstract: Specification-guided reinforcement learning (RL) provides a principled framework for encoding complex, temporally extended tasks using formal specificat

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Speech Enhancement Based on Drifting Models

DGX agent

arXiv:2604.24199v1 Announce Type: cross Abstract: We propose Speech Enhancement based on Drifting Models (DriftSE), a novel generative framework that formulates denoising as an equilibrium problem. Ra

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization

DGX agent

arXiv:2502.12672v4 Announce Type: replace-cross Abstract: Fine-tuning speech representation models can enhance performance on specific tasks but often compromises their cross-task generalization abili

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Sphere-Depth: A Benchmark for Depth Estimation Methods with Varying Spherical Camera Orientations

DGX agent

arXiv:2604.23432v1 Announce Type: cross Abstract: Reliable depth estimation from spherical images is crucial for 360{eg} vision in robotic navigation and immersive scene understanding. However, the on

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation

DGX agent

arXiv:2505.16637v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently demonstrated remarkable capabilities in machine translation (MT). However, most advanced MT-specifi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Stabilizing Efficient Reasoning with Step-Level Advantage Selection

DGX agent

arXiv:2604.24003v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong reasoning performance by allocating substantial computation at inference time, often generating long and ver

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning

DGX agent

arXiv:2604.23198v1 Announce Type: new Abstract: Current video moment retrieval excels at action-centric tasks but struggles with narrative content. Models can see extit{what is happening} but fail to

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Strategic Bidding in 6G Spectrum Auctions with Large Language Models

DGX agent

arXiv:2604.24156v1 Announce Type: cross Abstract: Efficient and fair spectrum allocation is a central challenge in 6G networks, where massive connectivity and heterogeneous services continuously compe

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

StratRAG: A Multi-Hop Retrieval Evaluation Dataset for Retrieval-Augmented Generation Systems

DGX agent

arXiv:2604.22757v1 Announce Type: cross Abstract: We introduce StratRAG, an open-source retrieval evaluation dataset for benchmarking Retrieval-Augmented Generation (RAG) systems on multi-hop reasonin

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Stress-Testing Emotional Support Models: Moving from Homogeneous to Diverse Help Seekers

DGX agent

arXiv:2601.07698v2 Announce Type: replace Abstract: As emotional support chatbots have recently gained significant traction across both research and industry, a common evaluation strategy has emerged:

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval

DGX agent

arXiv:2601.20597v2 Announce Type: replace Abstract: Continual Text-to-Video Retrieval (CTVR) is a challenging multimodal continual learning setting, where models must incrementally learn new semantic

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Structural Pruning of Large Vision Language Models: A Comprehensive Study on Pruning Dynamics, Recovery, and Data Efficiency

DGX agent

arXiv:2604.24380v1 Announce Type: new Abstract: While Large Vision Language Models (LVLMs) demonstrate impressive capabilities, their substantial computational and memory requirements pose deployment

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Supernodes and Halos: Loss-Critical Hubs in LLM Feed-Forward Layers

DGX agent

arXiv:2604.23475v1 Announce Type: cross Abstract: We study the organization of channel-level importance in transformer feed-forward networks (FFNs). Using a Fisher-style loss proxy (LP) based on activ

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SWE-QA: Can Language Models Answer Repository-level Code Questions?

DGX agent

arXiv:2509.14635v2 Announce Type: replace Abstract: Understanding and reasoning about entire software repositories is an essential capability for intelligent software engineering tools. While existing

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents

DGX agent

arXiv:2512.07538v3 Announce Type: replace Abstract: Recognizing semantic differences across documents is crucial for text generation evaluation and content alignment, especially in cross-lingual setti

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Switch Attention: Towards Dynamic and Fine-grained Hybrid Transformers

DGX agent

arXiv:2603.26380v2 Announce Type: replace Abstract: The attention mechanism has been the core component in modern transformer architectures. However, the computation of standard full attention scales

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SycoPhantasy: Quantifying Sycophancy and Hallucination in Small Open Weight VLMs for Vision-Language Scoring of Fantasy Characters

DGX agent

arXiv:2604.24346v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed as evaluators in tasks requiring nuanced image understanding, yet their reliability in scoring

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Symmetric Equilibrium Propagation for Thermodynamic Diffusion Training

DGX agent

arXiv:2604.23806v1 Announce Type: cross Abstract: The reverse process in score-based diffusion models is formally equivalent to overdamped Langevin dynamics in a time-dependent energy landscape. In ou

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SynthPert: Enhancing LLM Biological Reasoning via Synthetic Reasoning Traces for Cellular Perturbation Prediction

DGX agent

arXiv:2509.25346v2 Announce Type: replace Abstract: Predicting cellular responses to genetic perturbations represents a fundamental challenge in systems biology, critical for advancing therapeutic dis

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

TACO: Efficient Communication Compression of Intermediate Tensors for Scalable Tensor-Parallel LLM Training

DGX agent

arXiv:2604.24088v1 Announce Type: cross Abstract: Handling communication overhead in large-scale tensor-parallel training remains a critical challenge due to the dense, near-zero distributions of inte

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Test of Time: Rethinking Temporal Signal of Benchmark Contamination

DGX agent

arXiv:2509.00072v3 Announce Type: replace Abstract: Post-cutoff performance decay has been widely interpreted as a temporal signal for benchmark contamination. We critically examine this belief and de

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction

DGX agent

arXiv:2604.22880v1 Announce Type: new Abstract: Existing document OCR largely targets plain text or Markdown, discarding the structural and executable properties that make LaTeX essential for scientif

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering

DGX agent

arXiv:2604.24459v1 Announce Type: new Abstract: Despite recent advances in text-to-image generation, models still struggle to accurately render prompt-specified text with correct spatial layout -- esp

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers

DGX agent

arXiv:2604.24155v1 Announce Type: cross Abstract: The quest to align machine behavior with human values raises fundamental questions about the moral frameworks that should govern AI decision-making. M

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Kerimov-Alekberli Model: An Information-Geometric Framework for Real-Time System Stability

DGX agent

arXiv:2604.24083v1 Announce Type: new Abstract: This study introduces the Kerimov-Alekberli model, a novel information-geometric framework that redefines AI safety by formally linking non-equilibrium

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Optimal Sample Complexity of Multiclass and List Learning

DGX agent

arXiv:2604.24749v1 Announce Type: new Abstract: While the optimal sample complexity of binary classification in terms of the VC dimension is well-established, determining the optimal sample complexity

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

The Override Gap: A Magnitude Account of Knowledge Conflict Failure in Hypernetwork-Based Instant LLM Adaptation

DGX agent

arXiv:2604.23750v1 Announce Type: cross Abstract: Hypernetwork-based methods such as Doc-to-LoRA internalize a document into an LLM's weights in a single forward pass, but they fail systematically on

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Pragmatic Persona: Discovering LLM Persona through Bridging Inference

DGX agent

arXiv:2604.24079v1 Announce Type: cross Abstract: Large Language Models (LLMs) reveal inherent and distinctive personas through dialogue. However, most existing persona discovery approaches rely on su

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Price of Agreement: Measuring LLM Sycophancy in Agentic Financial Applications

DGX agent

arXiv:2604.24668v1 Announce Type: new Abstract: Given the increased use of LLMs in financial systems today, it becomes important to evaluate the safety and robustness of such systems. One failure mode

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Randomness Floor: Measuring Intrinsic Non-Randomness in Language Model Token Distributions

DGX agent

arXiv:2604.22771v1 Announce Type: cross Abstract: Language models cannot be random. This paper introduces Entropic Deviation (ED), the normalised KL divergence between a model's token distribution and

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Rise of Large Language Models and the Direction and Impact of US Federal Research Funding

DGX agent

arXiv:2601.15485v2 Announce Type: replace-cross Abstract: Federal research funding shapes the direction, diversity, and impact of the US scientific enterprise. Large language models (LLMs) are rapidly

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Security Cost of Intelligence: AI Capability, Cyber Risk, and Deployment Paradox

DGX agent

arXiv:2604.23058v1 Announce Type: cross Abstract: Firms are deploying more capable AI systems, but organizational controls often have not kept pace. These systems can generate greater productivity gai

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage

DGX agent

arXiv:2508.09603v2 Announce Type: replace Abstract: Membership inference attacks serves as useful tool for fair use of language models, such as detecting potential copyright infringement and auditing

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models

DGX agent

arXiv:2511.08577v2 Announce Type: replace-cross Abstract: Improving reasoning abilities of Large Language Models (LLMs), especially under parameter constraints, is crucial for real-world applications.

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

TOL: Textual Localization with OpenStreetMap

DGX agent

arXiv:2604.01644v2 Announce Type: replace Abstract: Natural language provides an intuitive way to express spatial intent in geospatial applications. While existing localization methods often rely on d

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

TopoHR: Hierarchical Centerline Representation for Cyclic Topology Reasoning in Driving Scenes with Point-to-Instance Relations

DGX agent

arXiv:2604.24119v1 Announce Type: new Abstract: Topology reasoning is crucial for autonomous driving. Current methods primarily focus on instance-level learning for centerline detection, followed by a

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Toward Polymorphic Backdoor against Semantic Communication via Intensity-Based Poisoning

DGX agent

arXiv:2604.23231v1 Announce Type: cross Abstract: Semantic Communication (SC) backdoor attacks aim to utilize triggers to manipulate the system into producing predetermined outputs via backdoored shar

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Towards Continual Expansion of Data Coverage: Automatic Text-guided Edge-case Synthesis

DGX agent

arXiv:2509.26158v2 Announce Type: replace-cross Abstract: The performance of deep neural networks is strongly influenced by the quality of their training data. However, mitigating dataset bias by manu

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills

DGX agent

arXiv:2603.25158v4 Announce Type: replace Abstract: Equipping Large Language Model (LLM) agents with domain-specific skills is critical for tackling complex tasks. Yet, manual authoring creates a seve

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Tube Diffusion Policy: Reactive Visual-Tactile Policy Learning for Contact-rich Manipulation

DGX agent

arXiv:2604.23609v1 Announce Type: new Abstract: Contact-rich manipulation is central to many everyday human activities, requiring continuous adaptation to contact uncertainty and external disturbances

model-releasesarxiv-cs-ro
28 Apr 2026
Model Releases

Ulterior Motives: Detecting Misaligned Reasoning in Continuous Thought Models

DGX agent

arXiv:2604.23460v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning has emerged as a key technique for eliciting complex reasoning in Large Language Models (LLMs). Although interpretable,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Understanding the Limits of Automated Evaluation for Code Review Bots in Practice

DGX agent

arXiv:2604.24525v1 Announce Type: cross Abstract: Automated code review (ACR) bots are increasingly used in industrial software development to assist developers during pull request (PR) review. As ado

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

UpstreamQA: A Modular Framework for Explicit Reasoning on Video Question Answering Tasks

DGX agent

arXiv:2604.23145v1 Announce Type: cross Abstract: Video Question Answering (VideoQA) demands models that jointly reason over spatial, temporal, and linguistic cues. However, the task's inherent comple

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

V-SEAM: Visual Semantic Editing and Attention Modulating for Causal Interpretability of Vision-Language Models

DGX agent

arXiv:2509.14837v2 Announce Type: replace Abstract: Recent advances in causal interpretability have extended from language models to vision-language models (VLMs), seeking to reveal their internal mec

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

VAPO: End-to-end Slide-Enhanced Speech Recognition with Omni-modal Large Language Models

DGX agent

arXiv:2510.08618v2 Announce Type: replace-cross Abstract: Omni-modal large language models (OLLMs) offer a promising end-to-end solution for slide-enhanced speech recognition due to their inherent mul

model-releasesarxiv-cs-cv
28 Apr 2026
← Previous
1…297298299300301…357
Next →