AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Meta-Inverse Physics-Informed Neural Networks for High-Dimensional Ordinary Differential Equations

DGX agent

arXiv:2605.03511v1 Announce Type: new Abstract: Solving inverse problems in dynamical systems governed by high-dimensional coupled ordinary differential equations (ODEs) is a ubiquitous challenge in s

model-releasesarxiv-cs-lg
6 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MHPR: Multidimensional Human Perception and Reasoning Benchmark for Large Vision-Languate Models

DGX agent

arXiv:2605.03485v1 Announce Type: new Abstract: Multidimensional human understanding is essential for real-world applications such as film analysis and virtual digital humans, yet current LVLM benchma

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

MILE: Mixture of Incremental LoRA Experts for Continual Semantic Segmentation across Domains and Modalities

DGX agent

arXiv:2605.03555v1 Announce Type: new Abstract: Continual semantic segmentation requires models to adapt to new domains or modalities without sacrificing performance on previously learned tasks. Exper

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Mitigating Frequency Learning Bias in Quantum Models via Multi-Stage Residual Learning

DGX agent

arXiv:2603.10083v2 Announce Type: replace-cross Abstract: Quantum machine learning models based on parameterized circuits can be viewed as Fourier series approximators. However, they often struggle to

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Mixed-Precision Information Bottlenecks for On-Device Trait-State Disentanglement in Bipolar Agitation Detection

DGX agent

arXiv:2605.03039v1 Announce Type: new Abstract: Continuous monitoring of bipolar disorder agitation via voice biomarkers requires disentangling stable speaker traits from volatile affective states on

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability

DGX agent

arXiv:2605.03217v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in settings that require nuanced ethical reasoning, yet existing bias evaluations treat model out

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Most ReLU Networks Admit Identifiable Parameters

DGX agent

arXiv:2605.03601v1 Announce Type: new Abstract: We study the realization map of deep ReLU networks, focusing on when a function determines its parameters up to scaling and permutation. To analyze hidd

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

MSEarth: A Multimodal Benchmark for Earth Science Phenomenon Discovery with MLLMs

DGX agent

arXiv:2505.20740v3 Announce Type: replace Abstract: The rapid advancement of multimodal large language models (MLLMs) offers new opportunities for complex scientific challenges, yet their application

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Multi-Agent Reasoning Improves Compute Efficiency: Pareto-Optimal Test-Time Scaling

DGX agent

arXiv:2605.01566v1 Announce Type: new Abstract: Advances in inference methods have enabled language models to improve their predictions without additional training. These methods often prioritize raw

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Multimodal Learning on Low-Quality Data with Conformal Predictive Self-Calibration

DGX agent

arXiv:2605.03820v1 Announce Type: new Abstract: Multimodal learning often grapples with the challenge of low-quality data, which predominantly manifests as two facets: modality imbalance and noisy cor

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Neuro-Symbolic Agents for Hallucination-Free Requirements Reuse

DGX agent

arXiv:2605.01562v1 Announce Type: cross Abstract: The Object-Oriented Method for Requirements Authoring and Management (OOMRAM) is a requirements reuse framework that relies on exact identifier matchi

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles

DGX agent

arXiv:2605.01847v1 Announce Type: new Abstract: Outcome-only evaluation under-specifies whether an evaluated agent profile preserves the commitments required to solve a multi-turn task coherently. Neu

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

New Bounds for Zarankiewicz Numbers via Reinforced LLM Evolutionary Search

DGX agent

arXiv:2605.01120v1 Announce Type: new Abstract: The Zarankiewicz number extbf{Z}(m, n, s, t) is the maximum number of edges in a bipartite graph G_{m, n} such that there is no complete K_{s, t} bipart

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Not that Groove: Zero-Shot Symbolic Music Editing

DGX agent

arXiv:2505.08203v2 Announce Type: replace-cross Abstract: While recent advancements in AI music generation have predominantly focused on direct audio synthesis, these systems suffer from inherent rigi

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

OCRR: A Benchmark for Online Correction Recovery under Distribution Shift

DGX agent

arXiv:2605.03153v1 Announce Type: cross Abstract: Static benchmarks measure a model frozen at training time. Real systems face distribution shift: new categories, paraphrased queries, drift: and must

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

On the Spectral Structure and Objective Equivalence of Orthogonal Multilabel Fisher Discriminants

DGX agent

arXiv:2605.03283v1 Announce Type: cross Abstract: We provide a unified theoretical analysis of Linear Discriminant Analysis with simultaneous multilabel scatter matrix formulations and Stiefel orthogo

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

On Verbalized Confidence Scores for LLMs

DGX agent

arXiv:2412.14737v2 Announce Type: replace Abstract: The rise of large language models (LLMs) and their tight integration into our daily life make it essential to dedicate efforts towards their trustwo

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Optimal control of the future via prospective learning with control

DGX agent

arXiv:2511.08717v4 Announce Type: replace-cross Abstract: Optimal control of the future is the next frontier for AI. Current approaches to this problem are typically rooted in reinforcement learning (

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

ORPilot: A Production-Oriented Agentic LLM-for-OR Tool for Optimization Modeling

DGX agent

arXiv:2605.02728v1 Announce Type: new Abstract: This paper presents ORPilot, an open-source agentic AI system that translates real-world business problems into solver-ready optimization models. Unlike

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Pairwise matrices for sparse autoencoders: single-feature inspection mislabels causal axes

DGX agent

arXiv:2605.03160v1 Announce Type: new Abstract: The standard sparse-autoencoder (SAE) interpretability protocol labels each feature from its top-activating contexts and validates by single-feature ste

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Parameter-Efficient Distributional RL via Normalizing Flows and a Geometry-Aware Cramer Surrogate

DGX agent

arXiv:2505.04310v2 Announce Type: replace-cross Abstract: Distributional Reinforcement Learning (DistRL) improves upon expectation-based methods by modeling full return distributions, but standard app

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Parameter-Efficient Multi-View Proficiency Estimation: From Discriminative Classification to Generative Feedback

DGX agent

arXiv:2605.03848v1 Announce Type: new Abstract: Estimating how well a person performs an action, rather than which action is performed, is central to coaching, rehabilitation, and talent identificatio

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

PatRe: A Full-Stage Office Action and Rebuttal Generation Benchmark for Patent Examination

DGX agent

arXiv:2605.03571v1 Announce Type: new Abstract: Patent examination is a complex, multi-stage process requiring both technical expertise and legal reasoning, increasingly challenged by rising applicati

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

PERSA: Reinforcement Learning for Professor-Style Personalized Feedback with LLMs

DGX agent

arXiv:2605.01123v1 Announce Type: new Abstract: Large language models (LLMs) can provide automated feedback in educational settings, but aligning an LLMs style with a specific instructors tone while m

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

PHBench: A Benchmark for Predicting Startup Series A Funding from Product Hunt Launch Signals

DGX agent

arXiv:2605.02974v1 Announce Type: cross Abstract: Structured launch signals on Product Hunt contain statistically significant predictive information for Series A funding outcomes. We construct PHBench

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments

DGX agent

arXiv:2605.02240v1 Announce Type: new Abstract: We introduce PhysicianBench, a benchmark for evaluating LLM agents on physician tasks grounded in real clinical setting within electronic health record

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

PIIGuard: Mitigating PII Harvesting under Adversarial Sanitization

DGX agent

arXiv:2605.03129v1 Announce Type: cross Abstract: Browsing-enabled LLM assistants can fetch webpages and answer contact-seeking queries, creating a practical channel for scraping contact-style persona

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

PODiff: Latent Diffusion in Proper Orthogonal Decomposition Space for Scientific Super-Resolution

DGX agent

arXiv:2605.03399v1 Announce Type: new Abstract: Probabilistic super-resolution of high-dimensional spatial fields using diffusion models is often computationally prohibitive due to the cost of operati

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

PriorNet: Prior-Guided Engagement Estimation from Face Video

DGX agent

arXiv:2605.03615v1 Announce Type: new Abstract: Engagement estimation from face video remains challenging because facial evidence is often incomplete, labeled data are limited, and engagement annotati

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Prism: Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diffusion Language Models

DGX agent

arXiv:2602.01842v3 Announce Type: replace Abstract: Inference-time compute has re-emerged as a practical way to improve LLM reasoning. Most test-time scaling (TTS) algorithms rely on autoregressive de

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

PROBE: Probabilistic Occupancy BEV Encoding with Analytical Translation Robustness for 3D Place Recognition

DGX agent

arXiv:2603.05965v2 Announce Type: replace-cross Abstract: We present PROBE (PRobabilistic Occupancy BEV Encoding), a learning-free LiDAR place recognition descriptor that models each BEV cell's occupa

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Quantum Hierarchical Reinforcement Learning via Variational Quantum Circuits

DGX agent

arXiv:2605.03434v1 Announce Type: new Abstract: Reinforcement learning is one of the most challenging learning paradigms where efficacy and efficiency gains are extremely valuable. Hierarchical reinfo

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

RAG over Thinking Traces Can Improve Reasoning Tasks

DGX agent

arXiv:2605.03344v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) has proven effective for knowledge-intensive tasks, but is widely believed to offer limited benefit for reasoning

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Raising the Ceiling: Better Empirical Fixation Densities for Saliency Benchmarking

DGX agent

arXiv:2605.03885v1 Announce Type: new Abstract: Empirical fixation densities, spatial distributions estimated from human eye-tracking data, are foundational to saliency benchmarking. They directly sha

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

RD-ViT: Recurrent-Depth Vision Transformer for Semantic Segmentation with Reduced Data Dependence Extending the Recurrent-Depth Transformer Architecture to Dense Prediction

DGX agent

arXiv:2605.03999v1 Announce Type: new Abstract: Vision Transformers (ViTs) achieve state-of-the-art segmentation accuracy but require large training datasets because each layer has unique parameters t

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Real Image Denoising with Knowledge Distillation for High-Performance Mobile NPUs

DGX agent

arXiv:2605.03680v1 Announce Type: new Abstract: While deep-learning-based image restoration has achieved unprecedented fidelity, deployment on mobile Neural Processing Units (NPUs) remains bottlenecke

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Reasoning Models Can be Accurately Pruned Via Chain-of-Thought Reconstruction

DGX agent

arXiv:2509.12464v2 Announce Type: replace Abstract: Reasoning language models such as DeepSeek-R1 produce long chain-of-thought traces during inference time which make them costly to deploy at scale.

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

ReCode: Reinforcing Code Generation with Reasoning-Process Rewards

DGX agent

arXiv:2508.05170v3 Announce Type: replace-cross Abstract: In practice, rigorous reasoning is often a key driver of correct code, while Reinforcement Learning (RL) for code generation often neglects op

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

ReLeaf: Benchmarking Leaf Segmentation across Domains and Species

DGX agent

arXiv:2605.03784v1 Announce Type: new Abstract: Rising global food demand and growing climate pressure increase the need for sustainable, precise agricultural practices. Automated, individualized plan

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Reproducing Complex Set-Compositional Information Retrieval

DGX agent

arXiv:2605.03824v1 Announce Type: new Abstract: Complex information needs may involve set-compositional queries using conjunction, disjunction, and exclusion, yet it remains unclear whether current re

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Rethinking Reasoning-Intensive Retrieval: Evaluating and Advancing Retrievers in Agentic Search Systems

DGX agent

arXiv:2605.04018v1 Announce Type: new Abstract: Reasoning-intensive retrieval aims to surface evidence that supports downstream reasoning rather than merely matching topical similarity. This capabilit

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Retrieval and Multi-Hop Reasoning in 1M-Token Context Windows: Evaluating LLMs on Classical Chinese Text

DGX agent

arXiv:2605.02173v1 Announce Type: new Abstract: We evaluate the long-context retrieval and reasoning capabilities of five frontier large language models with advertised 1M-token context windows on a c

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Retrieval-Augmented LLMs for Security Incident Analysis

DGX agent

arXiv:2603.18196v3 Announce Type: replace-cross Abstract: Investigating cybersecurity incidents requires collecting and analyzing evidence from multiple log sources, including intrusion detection aler

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Reward Hacking Benchmark: Measuring Exploits in LLM Agents with Tool Use

DGX agent

arXiv:2605.02964v1 Announce Type: new Abstract: Reinforcement learning (RL) trained language model agents with tool access are increasingly deployed in coding assistants, research tools, and autonomou

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

RFPrompt: Prompt-Based Expert Adaptation of the Large Wireless Model for Modulation Classification

DGX agent

arXiv:2605.03279v1 Announce Type: new Abstract: Automatic modulation classification (AMC) in real-world deployments demands robustness to distribution shifts arising from hardware impairments, unseen

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models

DGX agent

arXiv:2605.03821v1 Announce Type: new Abstract: Existing robot video world models are typically trained with low-level objectives such as reconstruction and perceptual similarity, which are poorly ali

model-releasesarxiv-cs-ro
6 May 2026
Model Releases

RoboEval: Where Robotic Manipulation Meets Structured and Scalable Evaluation

DGX agent

arXiv:2507.00435v2 Announce Type: replace-cross Abstract: We introduce RoboEval, a structured evaluation framework and benchmark for robotic manipulation that augments binary success with principled b

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Robust Language Identification for Romansh Varieties

DGX agent

arXiv:2603.15969v2 Announce Type: replace Abstract: The Romansh language has several regional varieties, called idioms, which sometimes have limited mutual intelligibility. Despite this linguistic div

model-releasesarxiv-cs-cl
6 May 2026
← Previous
1…274275276277278…361
Next →