AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

Kernel-Smith: A Unified Recipe for Evolutionary Kernel Optimization

DGX agent

arXiv:2603.28342v2 Announce Type: replace Abstract: We present Kernel-Smith, a framework for high-performance GPU kernel and operator generation that combines a stable evaluation-driven evolutionary a

model-releasesarxiv-cs-cl
24 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

KompeteAI: Accelerated Autonomous Multi-Agent System for End-to-End Pipeline Generation for Machine Learning Problems

DGX agent

arXiv:2508.10177v3 Announce Type: replace Abstract: Recent Large Language Model (LLM)-based AutoML systems demonstrate impressive capabilities but face significant limitations such as constrained expl

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Language as a Latent Variable for Reasoning Optimization

DGX agent

arXiv:2604.21593v1 Announce Type: new Abstract: As LLMs reduce English-centric bias, a surprising trend emerges: non-English responses sometimes outperform English on reasoning tasks. We hypothesize t

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Latent Denoising Improves Visual Alignment in Large Multimodal Models

DGX agent

arXiv:2604.21343v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) such as LLaVA are typically trained with an autoregressive language modeling objective, providing only indirect supervisi

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Learning to Communicate: Toward End-to-End Optimization of Multi-Agent Language Systems

DGX agent

arXiv:2604.21794v1 Announce Type: new Abstract: Multi-agent systems built on large language models have shown strong performance on complex reasoning tasks, yet most work focuses on agent roles and or

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Leveraging Multimodal LLMs for Built Environment and Housing Attribute Assessment from Street-View Imagery

DGX agent

arXiv:2604.21102v1 Announce Type: cross Abstract: We present a novel framework for automatically evaluating building conditions nationwide in the United States by leveraging large language models (LLM

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval

DGX agent

arXiv:2505.15269v2 Announce Type: replace Abstract: Recent developments in Video Large Language Models (Video LLMs) have enabled models to process hour-long videos and exhibit exceptional performance.

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Low-Rank Adaptation Redux for Large Models

DGX agent

arXiv:2604.21905v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) has emerged as the de facto standard for parameter-efficient fine-tuning (PEFT) of foundation models, enabling the adaptation

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

MaskDiME: Adaptive Masked Diffusion for Precise and Efficient Visual Counterfactual Explanations

DGX agent

arXiv:2602.18792v3 Announce Type: replace Abstract: Visual counterfactual explanations aim to reveal the minimal semantic modifications that can alter a model's prediction, providing causal and interp

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

MathDuels: Evaluating LLMs as Problem Posers and Solvers

DGX agent

arXiv:2604.21916v1 Announce Type: new Abstract: As frontier language models attain near-ceiling performance on static mathematical benchmarks, existing evaluations are increasingly unable to different

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

MATRAG: Multi-Agent Transparent Retrieval-Augmented Generation for Explainable Recommendations

DGX agent

arXiv:2604.20848v1 Announce Type: cross Abstract: Large Language Model (LLM)-based recommendation systems have demonstrated remarkable capabilities in understanding user preferences and generating per

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

MCAP: Deployment-Time Layer Profiling for Memory-Constrained LLM Inference

DGX agent

arXiv:2604.21026v1 Announce Type: new Abstract: Deploying large language models to heterogeneous hardware is often constrained by memory, not compute. We introduce MCAP (Monte Carlo Activation Profili

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Measuring Opinion Bias and Sycophancy via LLM-based Coercion

DGX agent

arXiv:2604.21564v1 Announce Type: new Abstract: Large language models increasingly shape the information people consume: they are embedded in search, consulted for professional advice, deployed as age

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

mGRADE: Minimal Recurrent Gating Meets Delay Convolutions for Lightweight Sequence Modeling

DGX agent

arXiv:2507.01829v2 Announce Type: replace-cross Abstract: Multi-timescale sequence modeling relies on capturing both local fast dynamics and global slow context; yet, maintaining these capabilities un

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

MISTY: High-Throughput Motion Planning via Mixer-based Single-step Drifting

DGX agent

arXiv:2604.21489v1 Announce Type: cross Abstract: Multi-modal trajectory generation is essential for safe autonomous driving, yet existing diffusion-based planners suffer from high inference latency d

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Multilingual and Domain-Agnostic Tip-of-the-Tongue Query Generation for Simulated Evaluation

DGX agent

arXiv:2604.21096v1 Announce Type: cross Abstract: Tip-of-the-Tongue (ToT) retrieval benchmarks have largely focused on English, limiting their applicability to multilingual information access. In this

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Musical Score Understanding Benchmark: Evaluating Large Language Models' Comprehension of Complete Musical Scores

DGX agent

arXiv:2511.20697v4 Announce Type: replace-cross Abstract: Understanding complete musical scores entails integrated reasoning over pitch, rhythm, harmony, and large-scale structure, yet the ability of

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Navigating the Clutter: Waypoint-Based Bi-Level Planning for Multi-Robot Systems

DGX agent

arXiv:2604.21138v1 Announce Type: cross Abstract: Multi-robot control in cluttered environments is a challenging problem that involves complex physical constraints, including robot-robot collisions, r

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Nemobot Games: Crafting Strategic AI Gaming Agents for Interactive Learning with Large Language Models

DGX agent

arXiv:2604.21896v1 Announce Type: new Abstract: This paper introduces a new paradigm for AI game programming, leveraging large language models (LLMs) to extend and operationalize Claude Shannon's taxo

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Neural surrogates for crystal growth dynamics with variable supersaturation: explicit vs. implicit conditioning

DGX agent

arXiv:2604.21753v1 Announce Type: cross Abstract: Simulations of crystal growth are performed by using Convolutional Recurrent Neural Network surrogate models, trained on a dataset of time sequences c

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

DGX agent

arXiv:2604.20911v1 Announce Type: cross Abstract: LLM agents deployed in production operate under operator-defined behavioral policies (system-prompt instructions such as prohibitions on credential di

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

OmniFit: Multi-modal 3D Body Fitting via Scale-agnostic Dense Landmark Prediction

DGX agent

arXiv:2604.21575v1 Announce Type: new Abstract: Fitting an underlying body model to 3D clothed human assets has been extensively studied, yet most approaches focus on either single-modal inputs such a

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

On the Role of Preprocessing and Memristor Dynamics in Reservoir Computing for Image Classification

DGX agent

arXiv:2604.21602v1 Announce Type: cross Abstract: Reservoir computing (RC) is an emerging recurrent neural network architecture that has attracted growing attention for its low training cost and modes

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Open-H-Embodiment: A Large-Scale Dataset for Enabling Foundation Models in Medical Robotics

DGX agent

arXiv:2604.21017v1 Announce Type: cross Abstract: Autonomous medical robots hold promise to improve patient outcomes, reduce provider workload, democratize access to care, and enable superhuman precis

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

OpenEstimate: Evaluating LLMs on Reasoning Under Uncertainty with Real-World Data

DGX agent

arXiv:2510.15096v2 Announce Type: replace Abstract: Real-world settings where language models (LMs) are deployed -- in domains spanning healthcare, finance, and other forms of knowledge work -- requir

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

OptiVerse: A Comprehensive Benchmark towards Optimization Problem Solving

DGX agent

arXiv:2604.21510v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning, complex optimization tasks remain challenging, requiring domain knowledge and robus

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Pretrain Where? Investigating How Pretraining Data Diversity Impacts Geospatial Foundation Model Performance

DGX agent

arXiv:2604.21104v1 Announce Type: new Abstract: New geospatial foundation models introduce a new model architecture and pretraining dataset, often sampled using different notions of data diversity. Pe

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

PREVENT-JACK: Context Steering for Swarms of Long Heavy Articulated Vehicles

DGX agent

arXiv:2604.21337v1 Announce Type: new Abstract: In this paper, we aim to extend the traditional point-mass-like robot representation in swarm robotics and instead study a swarm of long Heavy Articulat

model-releasesarxiv-cs-ro
24 Apr 2026
Model Releases

Principled Evaluation with Human Labels: One Rater at a Time and Rater Equivalence

DGX agent

arXiv:2106.01254v3 Announce Type: replace Abstract: In many classification tasks, there is no definitive ground truth, only human judgments that may disagree. We address two challenges that arise in s

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Process Supervision via Verbal Critique Improves Reasoning in Large Language Models

DGX agent

arXiv:2604.21611v1 Announce Type: cross Abstract: Inference-time scaling for LLM reasoning has focused on three axes: chain depth, sample breadth, and learned step-scorers (PRMs). We introduce a fourt

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

RailVQA: A Benchmark and Framework for Efficient Interpretable Visual Cognition in Automatic Train Operation

DGX agent

arXiv:2603.27112v2 Announce Type: replace Abstract: As Automatic Train Operation (ATO) advances toward GoA4 and beyond, it increasingly depends on efficient, reliable cab-view visual perception and de

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

RealRoute: Dynamic Query Routing System via Retrieve-then-Verify Paradigm

DGX agent

arXiv:2604.20860v1 Announce Type: cross Abstract: Despite the success of Retrieval-Augmented Generation (RAG) in grounding LLMs with external knowledge, its application over heterogeneous sources (e.g

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Reasoning About Traversability: Language-Guided Off-Road 3D Trajectory Planning

DGX agent

arXiv:2604.21249v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) enable high-level semantic reasoning for end-to-end autonomous driving, particularly in unstructured environments, e

model-releasesarxiv-cs-ro
24 Apr 2026
Model Releases

Rectified Schrodinger Bridge Matching for Few-Step Visual Navigation

DGX agent

arXiv:2604.05673v2 Announce Type: replace-cross Abstract: Visual navigation is a core challenge in Embodied AI, requiring autonomous agents to translate high-dimensional sensory observations into cont

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

ReFACT: A Benchmark for Scientific Confabulation Detection with Positional Error Annotations

DGX agent

arXiv:2509.25868v3 Announce Type: replace Abstract: The mechanisms underlying scientific confabulation in Large Language Models (LLMs) remain poorly understood. We introduce ReFACT (Reddit False And C

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Reinforcing 3D Understanding in Point-VLMs via Geometric Reward Credit Assignment

DGX agent

arXiv:2604.21160v1 Announce Type: new Abstract: Point-Vision-Language Models promise to empower embodied agents with executable spatial reasoning, yet they frequently succumb to geometric hallucinatio

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Reinforcing privacy reasoning in LLMs via normative simulacra from fiction

DGX agent

arXiv:2604.20904v1 Announce Type: cross Abstract: Information handling practices of LLM agents are broadly misaligned with the contextual privacy expectations of their users. Contextual Integrity (CI)

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Retrofit: Continual Learning with Controlled Forgetting for Binary Security Detection and Analysis

DGX agent

arXiv:2511.11439v2 Announce Type: replace-cross Abstract: Binary security has increasingly relied on deep learning to reason about malware behavior and program semantics. However, the performance ofte

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Revealing Geography-Driven Signals in Zone-Level Claim Frequency Models: An Empirical Study using Environmental and Visual Predictors

DGX agent

arXiv:2604.21893v1 Announce Type: cross Abstract: Geographic context is often consider relevant to motor insurance risk, yet public actuarial datasets provide limited location identifiers, constrainin

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

RewardBench 2: Advancing Reward Model Evaluation

DGX agent

arXiv:2506.01937v2 Announce Type: replace Abstract: Reward models are used throughout the post-training of language models to capture nuanced signals from preference data and provide a training target

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Robust Test-time Video-Text Retrieval: Benchmarking and Adapting for Query Shifts

DGX agent

arXiv:2604.20851v1 Announce Type: cross Abstract: Modern video-text retrieval (VTR) models excel on in-distribution benchmarks but are highly vulnerable to real-world query shifts, where the distribut

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

SatSAM2: Motion-Constrained Video Object Tracking in Satellite Imagery using Promptable SAM2 and Kalman Priors

DGX agent

arXiv:2511.18264v3 Announce Type: replace Abstract: Existing satellite video tracking methods often struggle with generalization, requiring scenario-specific training to achieve satisfactory performan

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Scaling of Gaussian Kolmogorov--Arnold Networks

DGX agent

arXiv:2604.21174v1 Announce Type: cross Abstract: The Gaussian scale parameter (epsilon) is central to the behavior of Gaussian Kolmogorov--Arnold Networks (KANs), yet its role in deep edge-based arch

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

SCASeg: Strip Cross-Attention for Efficient Semantic Segmentation

DGX agent

arXiv:2411.17061v2 Announce Type: replace Abstract: The Vision Transformer (ViT) has achieved notable success in computer vision, with its variants widely validated across various downstream tasks, in

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

SCM: Sleep-Consolidated Memory with Algorithmic Forgetting for Large Language Models

DGX agent

arXiv:2604.20943v1 Announce Type: new Abstract: We present SCM (Sleep-Consolidated Memory), a research preview of a memory architecture for large language models that draws on neuroscientific principl

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Secure LLM Fine-Tuning via Safety-Aware Probing

DGX agent

arXiv:2505.16737v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable success across many applications, but their ability to generate harmful content raises s

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models

DGX agent

arXiv:2604.21523v1 Announce Type: cross Abstract: Large Vision-Language Models (VLMs) are increasingly used to evaluate outputs of other models, for image-to-text (I2T) tasks such as visual question a

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Separable Expert Architecture: Toward Privacy-Preserving LLM Personalization via Composable Adapters and Deletable User Proxies

DGX agent

arXiv:2604.21571v1 Announce Type: new Abstract: Current model training approaches incorporate user information directly into shared weights, making individual data removal computationally infeasible w

model-releasesarxiv-cs-ai
24 Apr 2026
← Previous
1…302303304305306…357
Next →