AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

VGGT-Segmentor: Geometry-Enhanced Cross-View Segmentation

DGX agent

arXiv:2604.13596v1 Announce Type: new Abstract: Instance-level object segmentation across disparate egocentric and exocentric views is a fundamental challenge in visual understanding, critical for app

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body

DGX agent

arXiv:2512.14234v2 Announce Type: replace Abstract: Human communication is inherently multimodal and social: words, prosody, and body language jointly carry intent. Yet most prior systems model human

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?

DGX agent

arXiv:2503.23137v2 Announce Type: replace-cross Abstract: Understanding humor-particularly when it involves complex, contradictory narratives that require comparative reasoning-remains a significant c

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Working Notes on Late Interaction Dynamics: Analyzing Targeted Behaviors of Late Interaction Models

DGX agent

arXiv:2603.26259v2 Announce Type: replace-cross Abstract: While Late Interaction models exhibit strong retrieval performance, many of their underlying dynamics remain understudied, potentially hiding

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

WorkRB: A Community-Driven Evaluation Framework for AI in the Work Domain

DGX agent

arXiv:2604.13055v1 Announce Type: new Abstract: Today's evolving labor markets rely increasingly on recommender systems for hiring, talent management, and workforce analytics, with natural language pr

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

A Benchmark for Evaluating Outcome-Driven Constraint Violations in Autonomous AI Agents

DGX agent

arXiv:2512.20798v4 Announce Type: replace Abstract: As autonomous AI agents are deployed in high-stakes environments, ensuring their safety has become a paramount concern. Existing safety benchmarks p

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

A Foot Resistive Force Model for Legged Locomotion on Muddy Terrains

DGX agent

arXiv:2604.12006v1 Announce Type: new Abstract: Legged robots face significant challenges in moving and navigating on deformable and highly yielding terrain such as mud. We present a resistive force m

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

A General Model for Deepfake Speech Detection: Diverse Bonafide Resources or Diverse AI-Based Generators

DGX agent

arXiv:2603.27557v2 Announce Type: replace-cross Abstract: In this paper, we analyze two main factors of Bonafide Resource (BR) or AI-based Generator (AG) which affect the performance and the generalit

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

A Large-Scale Comparative Analysis of Imputation Methods for Single-Cell RNA Sequencing Data

DGX agent

arXiv:2603.24626v2 Announce Type: replace-cross Abstract: Background: Single-cell RNA sequencing (scRNA-seq) enables gene expression profiling at cellular resolution but is inherently affected by spar

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

A Layer-wise Analysis of Supervised Fine-Tuning

DGX agent

arXiv:2604.11838v1 Announce Type: cross Abstract: While critical for alignment, Supervised Fine-Tuning (SFT) incurs the risk of catastrophic forgetting, yet the layer-wise emergence of instruction-fol

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

A Sanity Check on Composed Image Retrieval

DGX agent

arXiv:2604.12904v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) aims to retrieve a target image based on a query composed of a reference image, and a relative caption that specifies the

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

DGX agent

arXiv:2602.11236v2 Announce Type: replace-cross Abstract: Building general-purpose embodied agents across diverse hardware remains a central challenge in robotics, often framed as the ''one-brain, man

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Adaptive Data Dropout: Towards Self-Regulated Learning in Deep Neural Networks

DGX agent

arXiv:2604.12945v1 Announce Type: cross Abstract: Deep neural networks are typically trained by uniformly sampling large datasets across epochs, despite evidence that not all samples contribute equall

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

AffectAgent: Collaborative Multi-Agent Reasoning for Retrieval-Augmented Multimodal Emotion Recognition

DGX agent

arXiv:2604.12735v1 Announce Type: new Abstract: LLM-based multimodal emotion recognition relies on static parametric memory and often hallucinates when interpreting nuanced affective states. In this p

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

AISafetyBenchExplorer: A Metric-Aware Catalogue of AI Safety Benchmarks Reveals Fragmented Measurement and Weak Benchmark Governance

DGX agent

arXiv:2604.12875v1 Announce Type: new Abstract: The rapid expansion of large language model (LLM) safety evaluation has produced a substantial benchmark ecosystem, but not a correspondingly coherent m

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

AlphaEval: Evaluating Agents in Production

DGX agent

arXiv:2604.12162v1 Announce Type: new Abstract: The rapid deployment of AI agents in commercial settings has outpaced the development of evaluation methodologies that reflect production realities. Exi

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Analyzing the Effect of Noise in LLM Fine-tuning

DGX agent

arXiv:2604.12469v1 Announce Type: new Abstract: Fine-tuning is the dominant paradigm for adapting pretrained large language models (LLMs) to downstream NLP tasks. In practice, fine-tuning datasets may

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

AnyPoC: Universal Proof-of-Concept Test Generation for Scalable LLM-Based Bug Detection

DGX agent

arXiv:2604.11950v1 Announce Type: cross Abstract: While recent LLM-based agents can identify many candidate bugs in source code, their reports remain static hypotheses that require manual validation,

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Are Video Reasoning Models Ready to Go Outside?

DGX agent

arXiv:2603.10652v2 Announce Type: replace-cross Abstract: In real-world deployment, vision-language models often encounter disturbances such as weather, occlusion, and camera motion. Under such condit

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

ARGOS: Who, Where, and When in Agentic Multi-Camera Person Search

DGX agent

arXiv:2604.12762v1 Announce Type: cross Abstract: We introduce ARGOS, the first benchmark and framework that reformulates multi-camera person search as an interactive reasoning problem requiring an ag

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

ASTRA: Let Arbitrary Subjects Transform in Video Editing

DGX agent

arXiv:2510.01186v2 Announce Type: replace Abstract: While existing video editing methods excel with single subjects, they struggle in dense, multi-subject scenes, frequently suffering from attention d

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Benchmarking Deflection and Hallucination in Large Vision-Language Models

DGX agent

arXiv:2604.12033v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) increasingly rely on retrieval to answer knowledge-intensive multimodal questions. Existing benchmarks overlook c

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Benchmarking Foundation Models with Retrieval-Augmented Generation in Olympic-Level Physics Problem Solving

DGX agent

arXiv:2510.00919v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) with foundation models has achieved strong performance across diverse tasks, but their capacity for exper

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Beyond Output Correctness: Benchmarking and Evaluating Large Language Model Reasoning in Coding Tasks

DGX agent

arXiv:2604.12379v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly rely on explicit reasoning to solve coding tasks, yet evaluating the quality of this reasoning remains chall

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Beyond Perception Errors: Semantic Fixation in Large Vision-Language Models

DGX agent

arXiv:2604.12119v1 Announce Type: new Abstract: Large vision-language models (VLMs) often rely on familiar semantic priors, but existing evaluations do not cleanly separate perception failures from ru

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Beyond Relevance: On the Relationship Between Retrieval and RAG Information Coverage

DGX agent

arXiv:2603.08819v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) systems combine document retrieval with a generative model to address complex information seeking tasks l

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Beyond Scores: Diagnostic LLM Evaluation via Fine-Grained Abilities

DGX agent

arXiv:2604.12191v1 Announce Type: new Abstract: Current evaluations of large language models aggregate performance across diverse tasks into single scores. This obscures fine-grained ability variation

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Beyond Single-Dimension Novelty: How Combinations of Theory, Method, and Results-based Novelty Shape Scientific Impact

DGX agent

arXiv:2604.12471v1 Announce Type: cross Abstract: Scientific novelty drives advances at the research frontier, yet it is also associated with heightened uncertainty and potential resistance from incum

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

BID-LoRA: A Parameter-Efficient Framework for Continual Learning and Unlearning

DGX agent

arXiv:2604.12686v1 Announce Type: cross Abstract: Recent advances in deep learning underscore the need for systems that can not only acquire new knowledge through Continual Learning (CL) but also remo

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Bilevel Late Acceptance Hill Climbing for the Electric Capacitated Vehicle Routing Problem

DGX agent

arXiv:2604.13013v1 Announce Type: new Abstract: This paper tackles the Electric Capacitated Vehicle Routing Problem (E-CVRP) through a bilevel optimization framework that handles routing and charging

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Bipedal-Walking-Dynamics Model on Granular Terrains

DGX agent

arXiv:2604.11981v1 Announce Type: new Abstract: Bipeds have demonstrated high agility and mobility in unstructured environments such as sand. The yielding of such granular media brings significant sin

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

Can AI Tools Transform Low-Demand Math Tasks? An Evaluation of Task Modification Capabilities

DGX agent

arXiv:2604.12743v1 Announce Type: new Abstract: While recent research has explored AI tools' ability to classify the quality of mathematical tasks (arXiv:2603.03512), little is known about their capac

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Can LLMs Beat Classical Hyperparameter Optimization Algorithms? A Study on autoresearch

DGX agent

arXiv:2603.24647v4 Announce Type: replace Abstract: The autoresearch repository enables an LLM agent to optimize hyperparameters by editing training code directly. We use it as a testbed to compare cl

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Chain-of-Models Pre-Training: Rethinking Training Acceleration of Vision Foundation Models

DGX agent

arXiv:2604.12391v1 Announce Type: cross Abstract: In this paper, we present Chain-of-Models Pre-Training (CoM-PT), a novel performance-lossless training acceleration method for vision foundation model

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Climate Model Tuning with Online Synchronization-Based Parameter Estimation

DGX agent

arXiv:2510.06180v2 Announce Type: replace-cross Abstract: In climate science, the tuning of climate models is a computationally intensive problem due to the combination of the high-dimensionality of t

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Clustering-Enhanced Domain Adaptation for Cross-Domain Intrusion Detection in Industrial Control Systems

DGX agent

arXiv:2604.12183v1 Announce Type: cross Abstract: Industrial control systems operate in dynamic environments where traffic distributions vary across scenarios, labeled samples are limited, and unknown

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Clustering with Uniformity- and Neighbor-Based Random Geometric Graphs

DGX agent

arXiv:2501.06268v3 Announce Type: replace Abstract: We propose a graph-based clustering method based on Cluster Catch Digraphs (CCDs) that extends their applicability to moderate-dimensional data sett

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

CoD-Lite: Real-Time Diffusion-Based Generative Image Compression

DGX agent

arXiv:2604.12525v1 Announce Type: new Abstract: Recent advanced diffusion methods typically derive strong generative priors by scaling diffusion transformers. However, scaling fails to generalize when

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

CoDe-R: Refining Decompiler Output with LLMs via Rationale Guidance and Adaptive Inference

DGX agent

arXiv:2604.12913v1 Announce Type: cross Abstract: Binary decompilation is a critical reverse engineering task aimed at reconstructing high-level source code from stripped executables. Although Large L

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation

DGX agent

arXiv:2604.12268v1 Announce Type: cross Abstract: Large language models (LLMs) can generate code from natural language, but the extent to which they capture intended program behavior remains unclear.

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

CODESTRUCT: Code Agents over Structured Action Spaces

DGX agent

arXiv:2604.05407v2 Announce Type: replace Abstract: LLM-based code agents treat repositories as unstructured text, applying edits through brittle string matching that frequently fails due to formattin

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Complementarity by Construction: A Lie-Group Approach to Solving Quadratic Programs with Linear Complementarity Constraints

DGX agent

arXiv:2604.11991v1 Announce Type: new Abstract: Many problems in robotics require reasoning over a mix of continuous dynamics and discrete events, such as making and breaking contact in manipulation a

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems

DGX agent

arXiv:2604.12312v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed as task-oriented agents in enterprise environments, ensuring their strict adherence to complex

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Cooperative Memory Paging with Keyword Bookmarks for Long-Horizon LLM Conversations

DGX agent

arXiv:2604.12376v1 Announce Type: cross Abstract: When LLM conversations grow beyond the context window, old content must be evicted -- but how does the model recover it when needed? We propose cooper

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

DecompSR: A dataset for decomposed analyses of compositional multihop spatial reasoning

DGX agent

arXiv:2511.02627v2 Announce Type: replace Abstract: We introduce DecompSR, decomposed spatial reasoning, a large benchmark dataset (over 5m datapoints) and generation framework designed to analyse com

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

DeCoNav: Dialog enhanced Long-Horizon Collaborative Vision-Language Navigation

DGX agent

arXiv:2604.12486v1 Announce Type: new Abstract: Long-horizon collaborative vision-language navigation (VLN) is critical for multi-robot systems to accomplish complex tasks beyond the capability of a s

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

Detecting Complex Money Laundering Patterns with Incremental and Distributed Graph Modeling

DGX agent

arXiv:2604.01315v2 Announce Type: replace Abstract: Money launderers take advantage of limitations in existing detection approaches by hiding their financial footprints in a deceitful manner. They man

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Disposition Distillation at Small Scale: A Three-Arc Negative Result

DGX agent

arXiv:2604.11867v1 Announce Type: cross Abstract: We set out to train behavioral dispositions (self-verification, uncertainty acknowledgment, feedback integration) into small language models (0.6B to

model-releasesarxiv-cs-ai
15 Apr 2026
← Previous
1…331332333334335…357
Next →