AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Model Releases

Beyond the Singular: Revealing the Value of Multiple Generations in Benchmark Evaluation

DGX agent

arXiv:2502.08943v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have demonstrated significant utility in real-world applications, exhibiting impressive capabilities in natural l

model-releasesarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CERSA: Cumulative Energy-Retaining Subspace Adaptation for Memory-Efficient Fine-Tuning

DGX agent

arXiv:2605.08174v1 Announce Type: cross Abstract: To mitigate the memory constraints associated with fine-tuning large pre-trained models, existing parameter-efficient fine-tuning (PEFT) methods, such

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Concordia: Self-Improving Synthetic Tables for Federated LLMs

DGX agent

arXiv:2605.09855v1 Announce Type: new Abstract: Federated learning (FL) enables training large language models (LLMs) without sharing raw data, but adapting LLMs under strict data isolation and non-II

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

ConQuR: Corner Aligned Activation Quantization via Optimized Rotations for LLMs

DGX agent

arXiv:2605.10793v1 Announce Type: new Abstract: Large language models (LLMs) are costly to deploy due to their large memory footprint and high inference cost. Weight-activation quantization can reduce

model-releasesarxiv-cs-lg
12 May 2026
Safety

Containment Verification: AI Safety Guarantees Independent of Alignment

DGX agent

arXiv:2605.09045v1 Announce Type: new Abstract: Agentic frameworks are the software layer through which AI agents act in the world. Existing safety methods intervene on the model and therefore remain

safetyarxiv-cs-ai
12 May 2026
Model Releases

Continuous Latent Contexts Enable Efficient Online Learning in Transformers

DGX agent

arXiv:2605.09867v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit a strong capacity for in-context learning: Given labeled examples, they can generate good predictions without par

model-releasesarxiv-cs-ai
12 May 2026
Research

DeepLevy: Learning Heavy-Tailed Uncertainty in Highly Volatile Time Series

DGX agent

arXiv:2605.10364v1 Announce Type: new Abstract: Modeling uncertainty in heavy-tailed time series remains a critical challenge for deep probabilistic forecasting models, which often struggle to capture

researcharxiv-cs-lg
12 May 2026
Model Releases

DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules

DGX agent

arXiv:2605.08614v1 Announce Type: new Abstract: Monitoring complex industrial assets relies on engineer-authored symbolic rules that trigger based on sensor conditions and prompt technicians to perfor

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Do Self-Evolving Agents Forget? Capability Degradation and Preservation in Lifelong LLM Agent Adaptation

DGX agent

arXiv:2605.09315v1 Announce Type: new Abstract: Recent advances in LLM agents enable systems that autonomously refine workflows, accumulate reusable skills, self-train their underlying models, and mai

model-releasesarxiv-cs-ai
12 May 2026
Research

Domain-Adaptive Arrhythmia Classification Using a Hybrid Transformer on Wearable Heart Signals

DGX agent

arXiv:2605.08199v1 Announce Type: cross Abstract: Cardiovascular disease remains the leading cause of death globally, underscoring the need for effective, accessible monitoring solutions, particularly

researcharxiv-cs-lg
12 May 2026
Model Releases

Drift is a Sampling Error: SNR-Aware Power Distributions for Long-Horizon Robotic Planning

DGX agent

arXiv:2605.09537v1 Announce Type: new Abstract: Despite rapid progress in Vision-Language-Action (VLA) models for robotic control, instruction drift remains a persistent failure mode in long-horizon t

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

DRNet: All-in-One Image Restoration via Prior-Guided Dynamic Reparameterization

DGX agent

arXiv:2605.08627v1 Announce Type: new Abstract: All-in-one image restoration aims to handle diverse degradations within a single model. However, existing methods often suffer from three key limitation

model-releasesarxiv-cs-cv
12 May 2026
Tutorials

EAM: Enhancing Anything with Diffusion Transformers for Blind Super-Resolution

DGX agent

arXiv:2505.05209v4 Announce Type: replace Abstract: Utilizing pre-trained Text-to-Image (T2I) diffusion models to guide Blind Super-Resolution (BSR) has become a predominant approach in the field. Whi

tutorialsarxiv-cs-cv
12 May 2026
Model Releases

Echo-LoRA: Parameter-Efficient Fine-Tuning via Cross-Layer Representation Injection

DGX agent

arXiv:2605.08177v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) has become a practical route for adapting large language models to downstream tasks, with LoRA-style methods be

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

EchoAlign: Bridging Generative and Discriminative Learning under Noisy Labels

DGX agent

arXiv:2405.12969v3 Announce Type: replace Abstract: Noisy labels severely hinder the accuracy and generalization of machine learning models, especially when ambiguous instance features make reliable a

model-releasesarxiv-cs-lg
12 May 2026
Applications

ERASE: Eliminating Redundant Visual Tokens via Adaptive Two-Stage Token Pruning

DGX agent

arXiv:2605.09982v1 Announce Type: new Abstract: Recent advancements in Vision-Language Models (VLMs) enable large language models (LLMs) to process high-resolution images, significantly improving real

applicationsarxiv-cs-cv
12 May 2026
Research

Explainable Knowledge Tracing via Probabilistic Embeddings and Pattern-based Reasoning

DGX agent

arXiv:2605.09369v1 Announce Type: new Abstract: Knowledge Tracing (KT) models students' knowledge states based on learning interactions to predict performance. While deep learning-based KT models have

researcharxiv-cs-ai
12 May 2026
Model Releases

Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition

DGX agent

arXiv:2605.10127v1 Announce Type: new Abstract: Recent research work on fashion outfit generation focuses on promoting visual consistency of garments by leveraging key information from reference image

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Function-Space ADMM for Decentralized Federated Learning: A Control Theoretic Perspective

DGX agent

arXiv:2605.09356v1 Announce Type: new Abstract: Decentralized federated learning (FL) is a promising approach for training machine learning models on sensor networks, Internet of Things (IoT) devices,

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Generative Actor-Critic with Soft Bridge Policies

DGX agent

arXiv:2605.08733v1 Announce Type: new Abstract: Expressive generative policies such as diffusion and flow models are appealing for MaxEnt online reinforcement learning because of their ability to mode

model-releasesarxiv-cs-lg
12 May 2026
Research

Geometry Guided Self-Consistency for Physical AI

DGX agent

arXiv:2605.08638v1 Announce Type: cross Abstract: State-of-the-art physical AI models generate a chunk of actions per inference through diffusion or flow matching, iteratively refining an initial nois

researcharxiv-cs-ai
12 May 2026
Model Releases

GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs

DGX agent

arXiv:2508.20325v3 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) become increasingly integral to various domains, their potential to generate harmful responses has prompted si

model-releasesarxiv-cs-ai
12 May 2026
Research

How Should LLMs Listen While Speaking? A Study of User-Stream Routing in Full-Duplex Spoken Dialogue

DGX agent

arXiv:2605.10199v1 Announce Type: new Abstract: Full-duplex spoken dialogue requires a model to keep listening while generating its own spoken response. This is challenging for large language models (

researcharxiv-cs-cl
12 May 2026
Model Releases

K12-KGraph: A Curriculum-Aligned Knowledge Graph for Benchmarking and Training Educational LLMs

DGX agent

arXiv:2605.09635v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in K-12 education, yet existing benchmarks such as C-Eval, CMMLU, GaokaoBench, and EduEval mainly eva

model-releasesarxiv-cs-cl
12 May 2026
Research

Kaczmarz Linear Attention

DGX agent

arXiv:2605.08587v1 Announce Type: cross Abstract: Long-context language modeling remains central to modern sequence modeling, but the quadratic cost of Transformer attention makes scaling computationa

researcharxiv-cs-ai
12 May 2026
Agents

LE-PAVD: Learning-Enhanced Physics-Aware Vehicle Dynamics for High-Speed Autonomous Navigation

DGX agent

arXiv:2605.08489v1 Announce Type: new Abstract: Accurate modeling of nonlinear vehicle dynamics is essential for high-speed autonomous racing, where controllers operate at the handling limits. Model-b

agentsarxiv-cs-ro
12 May 2026
Model Releases

LLaVA-UHD v4: What Makes Efficient Visual Encoding in MLLMs?

DGX agent

arXiv:2605.08985v1 Announce Type: new Abstract: Visual encoding constitutes a major computational bottleneck in Multimodal Large Language Models (MLLMs), especially for high-resolution image inputs. T

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

LLM Wardens: Mitigating Adversarial Persuasion with Third-Party Conversational Oversight

DGX agent

arXiv:2605.08321v1 Announce Type: cross Abstract: LLMs are increasingly capable of persuasion, which raises the question of how to protect users against manipulation. In a preregistered user study (N=

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Multi-Tier Labeling and Physics-Informed Learning for Orbital Anomaly Detection at Scale

DGX agent

arXiv:2605.09790v1 Announce Type: cross Abstract: Detecting orbital anomalies, such as maneuvers, atmospheric decay, and attitude upsets, across the rapidly growing population of low-Earth-orbit (LEO)

model-releasesarxiv-cs-ai
12 May 2026
Local Ai

On Distinguishing Capability Elicitation from Capability Creation in Post-Training: A Free-Energy Perspective

DGX agent

arXiv:2605.08368v1 Announce Type: new Abstract: Debates about large language model post-training often treat supervised fine-tuning (SFT) as imitation and reinforcement learning (RL) as discovery. But

local-aiarxiv-cs-ai
12 May 2026
Model Releases

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents

DGX agent

arXiv:2605.08876v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous agents that execute tool-augmented, multi-step tasks, where latency is a critical f

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Parameter-Efficient Neuroevolution for Diverse LLM Generation: Quality-Diversity Optimization via Prompt Embedding Evolution

DGX agent

arXiv:2605.09781v1 Announce Type: cross Abstract: Large Language Models exhibit mode collapse, producing homogeneous outputs that fail to explore valid solution spaces. We present QD-LLM, a framework

model-releasesarxiv-cs-ai
12 May 2026
Research

PLACO: A Multi-Stage Framework for Cost-Effective Performance in Human-AI Teams

DGX agent

arXiv:2605.08388v1 Announce Type: new Abstract: Human-AI teams play a pivotal role in improving overall system performance when neither the human nor the model can achieve such performance on their ow

researcharxiv-cs-ai
12 May 2026
Model Releases

PlantMarkerBench: A Multi-Species Benchmark for Evidence-Grounded Plant Marker Reasoning

DGX agent

arXiv:2605.10032v1 Announce Type: new Abstract: Cell-type-specific marker genes are fundamental to plant biology, yet existing resources primarily rely on curated databases or high-throughput studies

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

PrAg-PO: Prompt Augmented Policy Optimization for Robust and Diverse Mathematical Reasoning

DGX agent

arXiv:2602.03190v3 Announce Type: replace-cross Abstract: Reinforcement learning algorithms such as group-relative policy optimization (GRPO) have shown strong potential for improving the mathematical

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Preserving Foundational Capabilities in Flow-Matching VLAs through Conservative SFT

DGX agent

arXiv:2605.08879v1 Announce Type: new Abstract: Unconstrained fine-tuning of flow-matching Vision-Language-Action (VLA) models drives dense parameter overwrites, degrading pre-trained capabilities. We

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

Process Matters more than Output for Distinguishing Humans from Machines

DGX agent

arXiv:2605.06524v2 Announce Type: replace Abstract: Reliable human-machine discrimination is becoming increasingly important as large language models and autonomous agents are deployed in online setti

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Prompt-Activation Duality: Improving Activation Steering via Attention-Level Interventions

DGX agent

arXiv:2605.10664v1 Announce Type: cross Abstract: Activation steering controls language model behavior by adding directions to internal representations at inference time, but standard residual-stream

model-releasesarxiv-cs-ai
12 May 2026
Research

Provable Sparse Inversion and Token Relabel Enhanced One-shot Federated Learning with ViTs

DGX agent

arXiv:2605.10748v1 Announce Type: cross Abstract: One-Shot Federated Learning, where a central server learns a global model in a single communication round, has emerged as a promising paradigm. Howeve

researcharxiv-cs-ai
12 May 2026
Research

Quantum Transfer Learning Shows Improved Robustness in Low-Data Regimes

DGX agent

arXiv:2605.09118v1 Announce Type: cross Abstract: Transfer learning under limited data is a challenging setting, where models must adapt to new tasks with minimal supervision. Prior work has primarily

researcharxiv-cs-lg
12 May 2026
Model Releases

Real vs. Semi-Simulated: Rethinking Evaluation for Treatment Effect Estimation

DGX agent

arXiv:2605.10430v1 Announce Type: cross Abstract: Estimating heterogeneous treatment effects with machine learning has attracted substantial attention in both academic research and industrial practice

model-releasesarxiv-cs-ai
12 May 2026
Safety

Reasoning Compression with Mixed-Policy Distillation

DGX agent

arXiv:2605.08776v1 Announce Type: new Abstract: Reasoning-centric large language models (LLMs) achieve strong performance by generating intermediate reasoning trajectories, but often incur excessive t

safetyarxiv-cs-ai
12 May 2026
Model Releases

Robust Server Defense Against Unreliable Clients in One-Shot Fair Collaborative Machine Learning

DGX agent

arXiv:2605.08616v1 Announce Type: new Abstract: Collaborative machine learning (CML) enables multiple clients to train a global model jointly in a data-distributed setting. To address data privacy and

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

ROM: Real-time Overthinking Mitigation via Streaming Detection and Intervention

DGX agent

arXiv:2603.22016v2 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs) often reach a correct solution before their long Chain-of-Thought trace ends, yet continue with redundant verifi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Route Before Retrieve: Activating Latent Routing Abilities of LLMs for RAG vs. Long-Context Selection

DGX agent

arXiv:2605.10235v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have expanded the context window to beyond 128K tokens, enabling long-document understanding and multi-s

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

S2FT: Parameter-Efficient Fine-Tuning in Sparse Spectrum Domain

DGX agent

arXiv:2605.08589v1 Announce Type: new Abstract: Parameter Efficient Fine-Tuning (PEFT) is a key technique for adapting a large pretrained model to downstream tasks by fine-tuning only a small number o

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

SayNext-Bench: Why Do LLMs Struggle with Next-Utterance Anticipation?

DGX agent

arXiv:2602.00327v2 Announce Type: replace Abstract: We explore the use of large language models (LLMs) for next-utterance anticipation in human dialogue. Despite recent advances in LLMs demonstrating

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SciVQR: A Multidisciplinary Multimodal Benchmark for Advanced Scientific Reasoning Evaluation

DGX agent

arXiv:2605.10187v1 Announce Type: new Abstract: Scientific reasoning is a key aspect of human intelligence, requiring the integration of multimodal inputs, domain expertise, and multi-step inference a

model-releasesarxiv-cs-cv
12 May 2026
← Previous
1…469470471472473…1371
Next →