AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Caraman at SemEval-2026 Task 8: Three-Stage Multi-Turn Retrieval with Query Rewriting, Hybrid Search, and Cross-Encoder Reranking

DGX agent

arXiv:2605.12028v1 Announce Type: new Abstract: We describe our system for SemEval-2026 Task 8 (MTRAGEval), participating in Task A (Retrieval) across four English-language domains. Our approach emplo

model-releasesarxiv-cs-cl
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CATS: Cascaded Adaptive Tree Speculation for Memory-Limited LLM Inference Acceleration

DGX agent

arXiv:2605.11186v1 Announce Type: new Abstract: Auto-regressive decoding in Large Language Models (LLMs) is inherently memory-bound: every generation step requires loading the model weights and interm

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Checkup2Action: A Multimodal Clinical Check-up Report Dataset for Patient-Oriented Action Card Generation

DGX agent

arXiv:2605.11533v1 Announce Type: new Abstract: Clinical check-up reports are multimodal documents that combine page layouts, tables, numerical biomarkers, abnormality flags, imaging findings, and dom

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Chronicles-OCR: A Cross-Temporal Perception Benchmark for the Evolutionary Trajectory of Chinese Characters

DGX agent

arXiv:2605.11960v1 Announce Type: new Abstract: Vision Large Language Models (VLLMs) have achieved remarkable success in modern text-rich visual understanding. However, their perceptual robustness in

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

ClinicalBench: Stress-Testing Assertion-Aware Retrieval for Cross-Admission Clinical QA on MIMIC-IV

DGX agent

arXiv:2605.11143v1 Announce Type: new Abstract: Reasoning benchmarks measure clinical performance on clean inputs. We evaluate the step before reasoning: retrieval over real EHR notes, where negation,

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Control of Fully Actuated Aerial Vehicles: A Comparison of Model-based and Sensor-based Dynamic Inversion

DGX agent

arXiv:2605.12071v1 Announce Type: new Abstract: Fully actuated multirotor platforms decouple translational force generation from vehicle attitude, enabling independent control of position and orientat

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

CORE: Cyclic Orthotope Relation Embedding for Knowledge Graph Completion

DGX agent

arXiv:2605.11159v1 Announce Type: new Abstract: Knowledge graph completion (KGC) aims to automatically infer missing facts in multi-relational data by mapping entities and relations into continuous re

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Correcting Selection Bias in Sparse User Feedback for Large Language Model Quality Estimation: A Multi-Agent Hierarchical Bayesian Approach

DGX agent

arXiv:2605.12177v1 Announce Type: new Abstract: [Abridged] Production LLM deployments receive feedback from a non-random fraction of users: thumbs sit mostly in the tails of the satisfaction distribut

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Courtroom-Style Multi-Agent Debate with Progressive RAG and Role-Switching for Controversial Claim Verification

DGX agent

arXiv:2603.28488v2 Announce Type: replace Abstract: Large language models (LLMs) remain unreliable for high-stakes claim verification due to hallucinations and shallow reasoning. While retrieval-augme

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Covering Human Action Space for Computer Use: Data Synthesis and Benchmark

DGX agent

arXiv:2605.12501v1 Announce Type: new Abstract: Computer-use agents (CUAs) automate on-screen work, as illustrated by GPT-5.4 and Claude. Yet their reliability on complex, low-frequency interactions i

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Crash Assessment via Mesh-Based Graph Neural Networks and Physics-Aware Attention

DGX agent

arXiv:2605.11784v1 Announce Type: cross Abstract: Full-vehicle crash simulations are computationally expensive, limiting their use in iterative design exploration. This work investigates learned hybri

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

CTFusion: A CTF-based Benchmark for LLM Agent Evaluation

DGX agent

arXiv:2605.11504v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have enabled agentic systems for complex, multi-step tasks; cybersecurity is emerging as a prominent app

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

DarkQA: Benchmarking Vision-Language Models on Visual-Primitive Question Answering in Low-Light Indoor Scenes

DGX agent

arXiv:2512.24985v4 Announce Type: replace Abstract: Vision Language Models (VLMs) are increasingly adopted as central reasoning modules for embodied agents. Existing benchmarks evaluate their capabili

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Decomposing Evolutionary Mixture-of-LoRA Architectures: The Routing Lever, the Lifecycle Penalty, and a Substrate-Conditional Boundary

DGX agent

arXiv:2605.11153v1 Announce Type: new Abstract: We decompose an evolutionary mixture-of-LoRA system on a from-scratch ~150M-parameter widened-D substrate (D=1536, V=32000; D/V approx 0.048; the 'widen

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Delightful Gradients Accelerate Corner Escape

DGX agent

arXiv:2605.11908v1 Announce Type: new Abstract: Softmax policy gradient converges at O(1/t), but its transient behavior near sub-optimal corners of the simplex can be exponentially slow. The bottlenec

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Deploying Self-Supervised Learning for Real Seismic Data Denoising

DGX agent

arXiv:2605.11109v1 Announce Type: cross Abstract: Self-supervised learning (SSL) has emerged as a promising approach to seismic data denoising as it does not require clean reference data. In this work

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Detecting Data Contamination in LLMs via In-Context Learning

DGX agent

arXiv:2510.27055v2 Announce Type: replace Abstract: We present Contamination Detection via Context (CoDeC), a practical and accurate method to detect and quantify training data contamination in large

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

DiFaReli++: Diffusion Face Relighting with Consistent Cast Shadows

DGX agent

arXiv:2304.09479v5 Announce Type: replace Abstract: We introduce a novel approach to single-view face relighting in the wild, addressing challenges such as global illumination and cast shadows. A comm

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

DiffScore: Text Evaluation Beyond Autoregressive Likelihood

DGX agent

arXiv:2605.11601v1 Announce Type: new Abstract: Autoregressive language models are widely used for text evaluation, however, their left-to-right factorization introduces positional bias, i.e., early t

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

DisagMoE: Computation-Communication overlapped MoE Training via Disaggregated AF-Pipe Parallelism

DGX agent

arXiv:2605.11005v1 Announce Type: new Abstract: Mixture-of-experts (MoE) architectures enable trillion-parameter LLMs with sparsely activated experts. Expert parallelism (EP) is a widely adopted MoE t

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Distributed Pose Graph Optimization via Continuous Riemannian Dynamics

DGX agent

arXiv:2605.11210v1 Announce Type: new Abstract: We present a framework for distributed Pose Graph Optimization (PGO) by formulating the problem as a second-order continuous-time dynamical system evolv

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

Do Enterprise Systems Need Learned World Models? The Importance of Context to Infer Dynamics

DGX agent

arXiv:2605.12178v1 Announce Type: cross Abstract: World models enable agents to anticipate the effects of their actions by internalizing environment dynamics. In enterprise systems, however, these dyn

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Drop the Act: Probe-Filtered RL for Faithful Chain-of-Thought Reasoning

DGX agent

arXiv:2605.11467v1 Announce Type: new Abstract: Reasoning models post-hoc rationalize answers they have already committed to internally, producing chains of *reasoning theater*: deliberative-looking s

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback

DGX agent

arXiv:2506.13163v3 Announce Type: replace Abstract: We study the Logistic Contextual Slate Bandit problem, where, at each round, an agent selects a slate of N items from an exponentially large set (of

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Efficient and Adaptive Human Activity Recognition via LLM Backbones

DGX agent

arXiv:2605.12019v1 Announce Type: new Abstract: Human Activity Recognition (HAR) is a core task in pervasive computing systems, where models must operate under strict computational constraints while r

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Efficient LLM Reasoning via Variational Posterior Guidance with Efficiency Awareness

DGX agent

arXiv:2605.11019v1 Announce Type: new Abstract: Although large language models rely on chain-of-thought for complex reasoning, the overthinking phenomenon severely degrades inference efficiency. Exist

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

EgoEV-HandPose: Egocentric 3D Hand Pose Estimation and Gesture Recognition with Stereo Event Cameras

DGX agent

arXiv:2605.12297v1 Announce Type: new Abstract: Egocentric 3D hand pose estimation and gesture recognition are essential for immersive augmented/virtual reality, human-computer interaction, and roboti

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

EsoLang-Bench: Evaluating Genuine Reasoning in Large Language Models via Esoteric Programming Languages

DGX agent

arXiv:2603.09678v2 Announce Type: replace-cross Abstract: Large language models achieve near-ceiling performance on code generation benchmarks, yet most of the programming languages used by popular be

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Evaluating the Pre-Consultation Ability of LLMs using Diagnostic Guidelines

DGX agent

arXiv:2601.03627v3 Announce Type: replace Abstract: We introduce EPAG, a benchmark dataset and framework designed for Evaluating the Pre-consultation Ability of LLMs using diagnostic Guidelines. LLMs

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Exact Stiefel Optimization for Probabilistic PLS: Closed-Form Updates, Error Bounds, and Calibrated Uncertainty

DGX agent

arXiv:2605.11607v1 Announce Type: cross Abstract: Probabilistic partial least squares (PPLS) is a central likelihood-based model for two-view learning when one needs both interpretable latent factors

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks?

DGX agent

arXiv:2605.11086v1 Announce Type: cross Abstract: AI agents are rapidly gaining capabilities that could significantly reshape cybersecurity, making rigorous evaluation urgent. A critical capability is

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Extending Kernel Trick to Influence Functions

DGX agent

arXiv:2605.11239v1 Announce Type: new Abstract: In this paper, we present a dual representation of the influence functions, whose computational complexity scales with dataset size rather than model si

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

FastUMAP: Scalable Dimensionality Reduction via Bipartite Landmark Sampling

DGX agent

arXiv:2605.11428v1 Announce Type: new Abstract: Exploratory analysis of high-dimensional data rarely stops at a single embedding. In practice, analysts rerun dimensionality reduction after changing pr

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

fg-expo: Frontier-guided exploration-prioritized policy optimization via adaptive kl and gaussian curriculum

DGX agent

arXiv:2605.11403v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become the standard paradigm for LLM mathematical reasoning, with Group Relative Policy Opti

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Fine-Tuning Large Language Models for Cooperative Tactical Deconfliction of Small Unmanned Aerial Systems

DGX agent

arXiv:2603.28561v2 Announce Type: replace Abstract: The growing deployment of small Unmanned Aerial Systems (sUASs) in low-altitude airspaces has increased the need for reliable tactical deconfliction

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

Finite Volume-Informed Neural Network Framework for 2D Shallow Water Equations: Rugged Loss Landscapes and the Importance of Data Guidance

DGX agent

arXiv:2605.11001v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) are a simple surrogate-modelling paradigm for partial differential equations, but their standard strong-form re

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

FLAME: A New Dataset on FLemish Accounts of Momentary Experiences

DGX agent

arXiv:2504.14707v3 Announce Type: replace Abstract: We introduce FLAME (FLemish Accounts of Momentary Experiences), a new corpus of nearly 25,000 daily personal narratives in Belgian-Dutch (Flemish),

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

FLARE: Adaptive Multi-Dimensional Reputation for Robust Client Reliability in Federated Learning

DGX agent

arXiv:2511.14715v3 Announce Type: replace Abstract: Federated learning (FL) enables collaborative model training while preserving data privacy. However, it remains vulnerable to malicious clients who

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Focusable Monocular Depth Estimation

DGX agent

arXiv:2605.11756v1 Announce Type: new Abstract: Monocular depth foundation models generalize well across scenes, yet they are typically optimized with uniform pixel-wise objectives that do not disting

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Freeze Deep, Train Shallow: Interpretable Layer Allocation for Continued Pre-Training

DGX agent

arXiv:2605.11416v1 Announce Type: new Abstract: Selective layer-wise updates are essential for low-cost continued pre-training of Large Language Models (LLMs), yet determining which layers to freeze o

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

From Web to Pixels: Bringing Agentic Search into Visual Perception

DGX agent

arXiv:2605.12497v1 Announce Type: new Abstract: Visual perception connects high-level semantic understanding to pixel-level perception, but most existing settings assume that the decisive evidence for

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Generative Diffusion Prior Distillation for Long-Context Knowledge Transfer

DGX agent

arXiv:2605.11414v1 Announce Type: new Abstract: While traditional time-series classifiers assume full sequences at inference, practical constraints (latency and cost) often limit inputs to partial pre

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

GeneZip: Region-Aware Compression for Long Context DNA Modeling

DGX agent

arXiv:2602.17739v3 Announce Type: replace-cross Abstract: Long-context DNA models are limited by token-mixing cost and by how compression allocates representational budget across the genome. Existing

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Geometric Factual Recall in Transformers

DGX agent

arXiv:2605.12426v1 Announce Type: new Abstract: How do transformer language models memorize factual associations? A common view casts internal weight matrices as associative memories over pairs of emb

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

GeomHerd: A Forward-looking Herding Quantification via Ricci Flow Geometry on Agent Interactive Simulations

DGX agent

arXiv:2605.11645v1 Announce Type: cross Abstract: Herding -- where agents align their behaviors and act collectively -- is a central driver of market fragility and systemic risk. Existing approaches t

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

GeoR-Bench: Evaluating Geoscience Visual Reasoning

DGX agent

arXiv:2605.11541v1 Announce Type: new Abstract: Geoscience intelligence is expected to understand, reason about, and predict earth system changes to support human decision-making in critical domains s

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

GKnow: Measuring the Entanglement of Gender Bias and Factual Gender

DGX agent

arXiv:2605.12299v1 Announce Type: new Abstract: Recent works have analyzed the impact of individual components of neural networks on gendered predictions, often with a focus on mitigating gender bias.

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Gradient-Boosted Decision Tree for Listwise Context Model in Multimodal Review Helpfulness Prediction

DGX agent

arXiv:2305.12678v3 Announce Type: replace Abstract: Multimodal Review Helpfulness Prediction (MRHP) aims to rank product reviews based on predicted helpfulness scores and has been widely applied in e-

model-releasesarxiv-cs-cl
13 May 2026
← Previous
1…246247248249250…361
Next →