AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Epistemic Bias Injection: Manipulating LLM Opinion via Selective Context Retrieval

DGX agent

arXiv:2512.00804v3 Announce Type: replace-cross Abstract: When answering user queries, LLMs often retrieve knowledge from external sources stored in retrieval-augmented generation (RAG) databases. The

model-releasesarxiv-cs-ai
25 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

EPTS: Elastic Post-Training Sparsity for Efficient Large Language Model Compression

DGX agent

arXiv:2606.25285v1 Announce Type: new Abstract: Post-Training Sparsity (PTS) has emerged as a crucial paradigm for compressing Large Language Models to facilitate efficient deployment on resource-cons

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Evaluating LLMs on Real-World Software Performance Optimization

DGX agent

arXiv:2606.25530v1 Announce Type: cross Abstract: Software performance optimization is a notoriously complex and manual task. Despite the growing use of Large Language Models (LLMs) for code refinemen

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

EveLoad: Cognitive Workload Recognition from Event-Based Eye Movements

DGX agent

arXiv:2606.25177v1 Announce Type: new Abstract: Cognitive workload monitoring is important for adaptive rehabilitation and assistive interfaces, where task difficulty, pacing, and feedback should be a

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Evidence for feature-specific error correction in LLMs

DGX agent

arXiv:2606.24964v1 Announce Type: new Abstract: Understanding the features of large language models (LLMs) is a central goal of interpretability. LLMs are commonly assumed to use superposition to repr

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Evidential Perfusion Physics-Informed Neural Networks with Residual Uncertainty Quantification

DGX agent

arXiv:2603.09359v2 Announce Type: replace Abstract: Physics-informed neural networks (PINNs) have shown promise in addressing the ill-posed deconvolution problem in computed tomography perfusion (CTP)

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Exploring Information Seeking Agent Consolidation

DGX agent

arXiv:2602.00585v2 Announce Type: replace Abstract: Information-seeking agents have emerged as a powerful paradigm for knowledge-intensive tasks, yet today's systems remain specialized for the open we

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Failure Modes of Large Language Models on Research-Level Mathematics: A Taxonomy and an Empirical Characterisation

DGX agent

arXiv:2606.24902v1 Announce Type: cross Abstract: The 'First Proof' benchmark [1] posed ten research-level mathematics questions to the strongest publicly available LLMs and found them consistently wr

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Falcon: Functional Assembly and Language for Compositional Reasoning in X-ray

DGX agent

arXiv:2606.25701v1 Announce Type: new Abstract: Conventional vision-language models are largely object-centric, focusing on detecting and describing individual entities. In safety-critical X-ray bagga

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

FAR-LIO: Enabling High-Speed Autonomy through Fast, Accurate, and Robust LiDAR-Inertial Odometry

DGX agent

arXiv:2606.26010v1 Announce Type: new Abstract: Robust and accurate odometry estimation is essential in modern robotics. In environments characterized by highly dynamic motion and sensor noise, odomet

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming

DGX agent

arXiv:2606.19887v2 Announce Type: replace-cross Abstract: Existing safety benchmarks target general adversarial scenarios but miss finance-specific risks. Financial LLMs face regulatory compliance vio

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Flexible Gravitational-Wave Parameter Estimation with Transformers

DGX agent

arXiv:2512.02968v2 Announce Type: replace-cross Abstract: Gravitational-wave data analysis relies on accurate and efficient methods to extract physical information from noisy detector signals, yet the

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

FlowID : Enhancing Forensic Identification with Latent Flow-Matching Models

DGX agent

arXiv:2603.29591v2 Announce Type: replace Abstract: Every day, many people die under violent circumstances, whether from crimes, war, migration, or climate disasters. Medico-legal and law enforcement

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

FreeStory: Training-Free Character Consistency for Free-Form Visual Storytelling

DGX agent

arXiv:2606.25079v1 Announce Type: new Abstract: Visual storytelling aims to generate image sequences that are both aligned with narrative prompts and consistent in character appearance across images.

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models

DGX agent

arXiv:2606.25391v1 Announce Type: cross Abstract: Recent Large Audio Language Models (LALMs) have achieved remarkable progress in audio perceptual tasks across individual acoustic layers, including sp

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

From Sparse and Imperfect 2D Anchors to Consistent 3D Gaussian Street Scenes: Support-Aware Appearance

DGX agent

arXiv:2606.26007v1 Announce Type: new Abstract: Image priors can synthesize target conditions for 3D Gaussian street scenes, but independently edited views do not define a coherent 3D target. Direct f

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

FunPiQ: A New Benchmark for Pixel-Level Quality Assessment in Fundus Images

DGX agent

arXiv:2606.25915v1 Announce Type: new Abstract: Color fundus photography (CFP) is the most common ophthalmic imaging modality for large-scale screening. However, it is highly susceptible to degradatio

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Gaussian Mean Field Variational Inference can Overestimate Predictive Variance

DGX agent

arXiv:2606.25745v1 Announce Type: cross Abstract: Mean Field Variational Inference (MFVI) is widely understood to underestimate posterior variance. By analysing conjugate Bayesian Linear Regression (B

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Generating Input Distributions for Explaining Portfolio Optimization Pipelines

DGX agent

arXiv:2606.25808v1 Announce Type: cross Abstract: We propose a predict-optimize-explain framework that uses gradient-based sample generation to interpret various portfolio models by identifying macroe

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Geometry-Aware Online Scheduling for LLM Serving: From Theoretical Bound to System Practice

DGX agent

arXiv:2606.22327v2 Announce Type: replace Abstract: The explosive demand for interactive Large Language Model serving has highlighted the management of the Key-Value cache's dynamic memory footprint a

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

GroundSet: A Cadastral-Grounded Dataset for Spatial Understanding with Vector Data

DGX agent

arXiv:2603.14609v2 Announce Type: replace Abstract: Precise spatial understanding in Earth Observation is essential for translating raw aerial imagery into actionable insights for critical application

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

HG-Bench: A Benchmark for Multi-Page Handwritten Answer-Region Grounding in Automated Homework Assessment

DGX agent

arXiv:2606.25491v1 Announce Type: new Abstract: Automated homework assessment depends not only on recognizing student answers, but also on accurately locating where each answer and each intermediate r

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Hierarchical Graph Learning for Calendar Spread Strategies in Commodity Futures Markets

DGX agent

arXiv:2606.25811v1 Announce Type: cross Abstract: Commodity futures can be represented hierarchically, with underlying assets at the upper level and individual futures contracts at the lower level. En

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Hierarchical Reinforcement Learning for Neural Network Compression (HiReLC): Pruning and Quantization

DGX agent

arXiv:2606.26002v1 Announce Type: new Abstract: We present HiReLC, a hierarchical ensemble-reinforcement learning framework for automated joint quantization and structured pruning of deep neural netwo

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Hitting a Moving Target: Test-Time Adaptation for AI Text Detection under Continual Distribution Shift

DGX agent

arXiv:2606.25152v1 Announce Type: new Abstract: Deployed approaches for AI text detection often rely on training-time access to labeled datasets of both human-written and AI-generated text. This appro

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

How Reliable Is Your Jailbreak Judge? Calibration and Adversarial Robustness of Automated ASR Scoring

DGX agent

arXiv:2606.25487v1 Announce Type: new Abstract: Almost every paper on LLM jailbreaks and prompt injection reports an attack-success rate (ASR), and that number is assigned not by people but by an auto

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations

DGX agent

arXiv:2606.26041v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance on OCR-based benchmarks and increasingly focused on text-rich understanding, but their

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

How Small Can 6G Reason? Scaling Tiny-to-Small Language Models for AI-Native Networks

DGX agent

arXiv:2603.02156v2 Announce Type: replace-cross Abstract: Emerging 6G visions, reflected in ongoing standardization efforts within 3GPP, IETF, ETSI, ITU-T, and the O-RAN Alliance, increasingly charact

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Improving Factuality of 3D Brain MRI Report Generation with Paired Image-domain Retrieval and Text-domain Augmentation

DGX agent

arXiv:2411.15490v2 Announce Type: replace Abstract: Acute ischemic stroke (AIS) requires time-critical decision-making, where inaccurate interpretation of neuroimaging findings can lead to irreversibl

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Improving Zero-Shot Offline RL via Behavioral Task Sampling

DGX agent

arXiv:2604.25496v2 Announce Type: replace Abstract: Offline zero-shot reinforcement learning (RL) aims to learn agents that optimize unseen reward functions without additional environment interaction.

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

In-Context World Modeling for Robotic Control

DGX agent

arXiv:2606.26025v1 Announce Type: cross Abstract: Modern Vision-Language-Action (VLA) models often fail to generalize to novel setups, such as altered camera viewpoints or robot morphologies, because

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

IndicContextEval: A Benchmark for Evaluating Context Utilisation in Audio Large Language Models Across 8 Indic Languages

DGX agent

arXiv:2606.19157v2 Announce Type: replace-cross Abstract: AudioLLMs enable speech recognition conditioned on textual prompts such as domain descriptions or entity lists. However, it remains unclear wh

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Internal Data Repetition Destroys Language Models

DGX agent

arXiv:2606.24998v1 Announce Type: new Abstract: Language models are running out of high-quality training data, and even aggressively deduplicated corpora retain some amount of repetition. Earlier cont

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Inverse Reinforcement Learning for Interpretable Keystroke Biomarkers in Parkinson's Disease

DGX agent

arXiv:2606.25270v1 Announce Type: new Abstract: Keystroke dynamics have been explored extensively as a passive digital biomarker for Parkinson's disease (PD), typically by extracting summary statistic

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

InvestPhilBench: A Multi-Layer Dynamic Benchmark for Evaluating Large Language Model Procedural Reasoning in Expert Investment Philosophy

DGX agent

arXiv:2606.25984v1 Announce Type: cross Abstract: Large language models are increasingly deployed as investment research assistants, yet no benchmark tests whether they can accurately reconstruct and

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Invoice Haystack: Benchmarking Document Retrieval and Visual Question Answering Under Strong Visual Homogeneity

DGX agent

arXiv:2606.25343v1 Announce Type: new Abstract: Vision Language Models have achieved near-human performance on single-document Visual Question Answering, yet their effectiveness degrades significantly

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

KidRisk: Benchmark Dataset for Children Dangerous Action Recognition

DGX agent

arXiv:2606.25298v1 Announce Type: new Abstract: Children are naturally energetic, and during their spontaneous activities, they often encounter potentially dangerous situations, especially when lackin

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Laplace--Fisher Gate Identities for Optimal Matrix-Gated Blended Score Estimation

DGX agent

arXiv:2606.25169v1 Announce Type: cross Abstract: Sampling from an unnormalized target by reversing an Ornstein--Uhlenbeck diffusion requires the score of each noise-perturbed marginal. Tweedie's iden

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Latent Block-Diffusion Temporal Point Processes: A Semi-Autoregressive Framework for Asynchronous Event Sequence Generation

DGX agent

arXiv:2606.24982v1 Announce Type: new Abstract: Modeling and sampling from the underlying distribution of asynchronous event sequences are crucial in various real-world applications, including social

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Learning Dynamical Systems from Multiple Sparse Datasets: A Hierarchical Bayesian Modeling Approach

DGX agent

arXiv:2606.24966v1 Announce Type: new Abstract: Estimating parameters of dynamical systems from sparse, noisy, and irregularly sampled data is often severely ill-conditioned. When multiple related dat

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Learning to Adapt: Reptile-D-Learning for Robust and Efficient Control Under Parametric Uncertainty

DGX agent

arXiv:2606.25659v1 Announce Type: new Abstract: Learning-based Lyapunov Control (LLC) provides formal stability guarantees for nonlinear systems, but its validity relies heavily on accurate system mod

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

LEVIRDet: A Million-Scale 159-Category Dataset and Foundation Model for Universal Remote Sensing Object Detection

DGX agent

arXiv:2606.25312v1 Announce Type: new Abstract: Remote sensing object detection has advanced rapidly with the development of large-scale benchmarks and modern detection architectures. However, existin

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

LibEvoBench: Probing Temporal Knowledge Stratification in Code Generation Models

DGX agent

arXiv:2606.25402v1 Announce Type: cross Abstract: Large software projects often depend on older versions of libraries, even as APIs continue to evolve across releases. This creates a challenge for LLM

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

LLM Performance on a Real, Double-Marked GCSE Benchmark

DGX agent

arXiv:2606.24973v1 Announce Type: new Abstract: We introduce a dataset of 32,534 double-marked real student responses to GCSE mock exams (GCSEs are the UK's national exams, taken at age ~16), spanning

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

MacroLens: A Multi-Task Benchmark for Contextual Financial Reasoning under Macroeconomic Scenarios

DGX agent

arXiv:2606.24950v1 Announce Type: new Abstract: Financial decision-making is contextual: forecasting prices, valuing companies, and assessing event exposure weigh price history, accounting fundamental

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

MedLayBench-V: A Large-Scale Benchmark for Expert-Lay Semantic Alignment in Medical Vision Language Models

DGX agent

arXiv:2604.05738v2 Announce Type: replace Abstract: Medical Vision-Language Models (Med-VLMs) have achieved expert-level proficiency in interpreting diagnostic imaging. However, current models are pre

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Memory-Efficient Policy Libraries with Low-Rank Adaptation in Reinforcement Learning

DGX agent

arXiv:2606.25700v1 Announce Type: new Abstract: When fine-tuning Large Language Models (LLMs), there has been success in minimizing both memory usage and computation with Parameter-Efficient Fine-Tuni

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

MINIF2F-DAFNY: LLM-Guided Mathematical Theorem Proving via Auto-Active Verification

DGX agent

arXiv:2512.10187v3 Announce Type: replace Abstract: LLMs excel at reasoning, but validating their steps remains challenging. Formal verification offers a solution through mechanically checkable proofs

model-releasesarxiv-cs-lg
25 Jun 2026
← Previous
1…118119120121122…361
Next →