AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Model Releases

AgentCollabBench: Diagnosing When Good Agents Make Bad Collaborators

DGX agent

arXiv:2605.08647v1 Announce Type: cross Abstract: Multi-agent systems achieve state-of-the-art outcomes through peer collaboration. However, when an agent in the pipeline silently drops a constraint,

model-releasesarxiv-cs-ai
12 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

AIPO: : Learning to Reason from Active Interaction

DGX agent

arXiv:2605.08401v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have demonstrated remarkable reasoning capabilities, largely stimulated by Reinforcement Learning with

safetyarxiv-cs-ai
12 May 2026
Safety

Aligning LLM Uncertainty with Human Disagreement in Subjectivity Analysis

DGX agent

arXiv:2605.10415v1 Announce Type: new Abstract: Large language models for subjectivity analysis are typically trained with aggregated labels, which compress variations in human judgment into a single

safetyarxiv-cs-cl
12 May 2026
Model Releases

Artificial Intelligence in Number Theory: LLMs for Algorithm Generation and Ensemble Methods for Conjecture Verification

DGX agent

arXiv:2504.19451v3 Announce Type: cross Abstract: This paper presents two concrete applications of Artificial Intelligence to algorithmic and analytic number theory. Recent benchmarks of large languag

model-releasesarxiv-cs-ai
12 May 2026
Safety

ASACK : Adaptive Safe Active Continual Koopman Learning for Uncertain Systems with Contractive Guarantees

DGX agent

arXiv:2605.09659v1 Announce Type: new Abstract: Koopman operator theory provides a powerful framework for representing nonlinear dynamics through a linear operator acting on lifted observables, enabli

safetyarxiv-cs-ro
12 May 2026
Safety

Auditing Data Membership in Reinforcement Learning With Verifiable Rewards

DGX agent

arXiv:2511.14045v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a core training stage in recent large language models (LLMs). Its reliance on

safetyarxiv-cs-ai
12 May 2026
Model Releases

Benchmarking Compositional Generalisation for Machine Learning Interatomic Potentials

DGX agent

arXiv:2605.08988v1 Announce Type: cross Abstract: Machine Learning Interatomic Potentials play a fundamental role in computational chemistry and materials science, enabling applications from molecular

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Benchmarking Transformer and xLSTM for Time-Series Forecasting of Heat Consumption

DGX agent

arXiv:2605.09722v1 Announce Type: new Abstract: Obtaining an accurate short-term forecasting for heat demand is an essential part of operating district heating networks cost-efficient and reliable. He

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Beyond the All-in-One Agent: Benchmarking Role-Specialized Multi-Agent Collaboration in Enterprise Workflows

DGX agent

arXiv:2605.08761v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly expected to operate in enterprise environments, where work is distributed across specialized roles,

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Beyond the False Trade-off: Adaptive EWC for Stealthy and Generalizable T2I Backdoors

DGX agent

arXiv:2605.08280v1 Announce Type: cross Abstract: Preserving model fidelity is essential for stealthy text-to-image (T2I) backdoor attacks. Existing methods such as Learning without Forgetting (LwF) r

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Beyond the Singular: Revealing the Value of Multiple Generations in Benchmark Evaluation

DGX agent

arXiv:2502.08943v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have demonstrated significant utility in real-world applications, exhibiting impressive capabilities in natural l

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CERSA: Cumulative Energy-Retaining Subspace Adaptation for Memory-Efficient Fine-Tuning

DGX agent

arXiv:2605.08174v1 Announce Type: cross Abstract: To mitigate the memory constraints associated with fine-tuning large pre-trained models, existing parameter-efficient fine-tuning (PEFT) methods, such

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Concordia: Self-Improving Synthetic Tables for Federated LLMs

DGX agent

arXiv:2605.09855v1 Announce Type: new Abstract: Federated learning (FL) enables training large language models (LLMs) without sharing raw data, but adapting LLMs under strict data isolation and non-II

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

ConQuR: Corner Aligned Activation Quantization via Optimized Rotations for LLMs

DGX agent

arXiv:2605.10793v1 Announce Type: new Abstract: Large language models (LLMs) are costly to deploy due to their large memory footprint and high inference cost. Weight-activation quantization can reduce

model-releasesarxiv-cs-lg
12 May 2026
Safety

Containment Verification: AI Safety Guarantees Independent of Alignment

DGX agent

arXiv:2605.09045v1 Announce Type: new Abstract: Agentic frameworks are the software layer through which AI agents act in the world. Existing safety methods intervene on the model and therefore remain

safetyarxiv-cs-ai
12 May 2026
Model Releases

Continuous Latent Contexts Enable Efficient Online Learning in Transformers

DGX agent

arXiv:2605.09867v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit a strong capacity for in-context learning: Given labeled examples, they can generate good predictions without par

model-releasesarxiv-cs-ai
12 May 2026
Research

DeepLevy: Learning Heavy-Tailed Uncertainty in Highly Volatile Time Series

DGX agent

arXiv:2605.10364v1 Announce Type: new Abstract: Modeling uncertainty in heavy-tailed time series remains a critical challenge for deep probabilistic forecasting models, which often struggle to capture

researcharxiv-cs-lg
12 May 2026
Model Releases

DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules

DGX agent

arXiv:2605.08614v1 Announce Type: new Abstract: Monitoring complex industrial assets relies on engineer-authored symbolic rules that trigger based on sensor conditions and prompt technicians to perfor

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Do Self-Evolving Agents Forget? Capability Degradation and Preservation in Lifelong LLM Agent Adaptation

DGX agent

arXiv:2605.09315v1 Announce Type: new Abstract: Recent advances in LLM agents enable systems that autonomously refine workflows, accumulate reusable skills, self-train their underlying models, and mai

model-releasesarxiv-cs-ai
12 May 2026
Research

Domain-Adaptive Arrhythmia Classification Using a Hybrid Transformer on Wearable Heart Signals

DGX agent

arXiv:2605.08199v1 Announce Type: cross Abstract: Cardiovascular disease remains the leading cause of death globally, underscoring the need for effective, accessible monitoring solutions, particularly

researcharxiv-cs-lg
12 May 2026
Model Releases

Drift is a Sampling Error: SNR-Aware Power Distributions for Long-Horizon Robotic Planning

DGX agent

arXiv:2605.09537v1 Announce Type: new Abstract: Despite rapid progress in Vision-Language-Action (VLA) models for robotic control, instruction drift remains a persistent failure mode in long-horizon t

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

DRNet: All-in-One Image Restoration via Prior-Guided Dynamic Reparameterization

DGX agent

arXiv:2605.08627v1 Announce Type: new Abstract: All-in-one image restoration aims to handle diverse degradations within a single model. However, existing methods often suffer from three key limitation

model-releasesarxiv-cs-cv
12 May 2026
Tutorials

EAM: Enhancing Anything with Diffusion Transformers for Blind Super-Resolution

DGX agent

arXiv:2505.05209v4 Announce Type: replace Abstract: Utilizing pre-trained Text-to-Image (T2I) diffusion models to guide Blind Super-Resolution (BSR) has become a predominant approach in the field. Whi

tutorialsarxiv-cs-cv
12 May 2026
Model Releases

Echo-LoRA: Parameter-Efficient Fine-Tuning via Cross-Layer Representation Injection

DGX agent

arXiv:2605.08177v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) has become a practical route for adapting large language models to downstream tasks, with LoRA-style methods be

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

EchoAlign: Bridging Generative and Discriminative Learning under Noisy Labels

DGX agent

arXiv:2405.12969v3 Announce Type: replace Abstract: Noisy labels severely hinder the accuracy and generalization of machine learning models, especially when ambiguous instance features make reliable a

model-releasesarxiv-cs-lg
12 May 2026
Applications

ERASE: Eliminating Redundant Visual Tokens via Adaptive Two-Stage Token Pruning

DGX agent

arXiv:2605.09982v1 Announce Type: new Abstract: Recent advancements in Vision-Language Models (VLMs) enable large language models (LLMs) to process high-resolution images, significantly improving real

applicationsarxiv-cs-cv
12 May 2026
Research

Explainable Knowledge Tracing via Probabilistic Embeddings and Pattern-based Reasoning

DGX agent

arXiv:2605.09369v1 Announce Type: new Abstract: Knowledge Tracing (KT) models students' knowledge states based on learning interactions to predict performance. While deep learning-based KT models have

researcharxiv-cs-ai
12 May 2026
Model Releases

Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition

DGX agent

arXiv:2605.10127v1 Announce Type: new Abstract: Recent research work on fashion outfit generation focuses on promoting visual consistency of garments by leveraging key information from reference image

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Function-Space ADMM for Decentralized Federated Learning: A Control Theoretic Perspective

DGX agent

arXiv:2605.09356v1 Announce Type: new Abstract: Decentralized federated learning (FL) is a promising approach for training machine learning models on sensor networks, Internet of Things (IoT) devices,

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Generative Actor-Critic with Soft Bridge Policies

DGX agent

arXiv:2605.08733v1 Announce Type: new Abstract: Expressive generative policies such as diffusion and flow models are appealing for MaxEnt online reinforcement learning because of their ability to mode

model-releasesarxiv-cs-lg
12 May 2026
Research

Geometry Guided Self-Consistency for Physical AI

DGX agent

arXiv:2605.08638v1 Announce Type: cross Abstract: State-of-the-art physical AI models generate a chunk of actions per inference through diffusion or flow matching, iteratively refining an initial nois

researcharxiv-cs-ai
12 May 2026
Model Releases

GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs

DGX agent

arXiv:2508.20325v3 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) become increasingly integral to various domains, their potential to generate harmful responses has prompted si

model-releasesarxiv-cs-ai
12 May 2026
Research

How Should LLMs Listen While Speaking? A Study of User-Stream Routing in Full-Duplex Spoken Dialogue

DGX agent

arXiv:2605.10199v1 Announce Type: new Abstract: Full-duplex spoken dialogue requires a model to keep listening while generating its own spoken response. This is challenging for large language models (

researcharxiv-cs-cl
12 May 2026
Model Releases

K12-KGraph: A Curriculum-Aligned Knowledge Graph for Benchmarking and Training Educational LLMs

DGX agent

arXiv:2605.09635v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in K-12 education, yet existing benchmarks such as C-Eval, CMMLU, GaokaoBench, and EduEval mainly eva

model-releasesarxiv-cs-cl
12 May 2026
Research

Kaczmarz Linear Attention

DGX agent

arXiv:2605.08587v1 Announce Type: cross Abstract: Long-context language modeling remains central to modern sequence modeling, but the quadratic cost of Transformer attention makes scaling computationa

researcharxiv-cs-ai
12 May 2026
Agents

LE-PAVD: Learning-Enhanced Physics-Aware Vehicle Dynamics for High-Speed Autonomous Navigation

DGX agent

arXiv:2605.08489v1 Announce Type: new Abstract: Accurate modeling of nonlinear vehicle dynamics is essential for high-speed autonomous racing, where controllers operate at the handling limits. Model-b

agentsarxiv-cs-ro
12 May 2026
Model Releases

LLaVA-UHD v4: What Makes Efficient Visual Encoding in MLLMs?

DGX agent

arXiv:2605.08985v1 Announce Type: new Abstract: Visual encoding constitutes a major computational bottleneck in Multimodal Large Language Models (MLLMs), especially for high-resolution image inputs. T

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

LLM Wardens: Mitigating Adversarial Persuasion with Third-Party Conversational Oversight

DGX agent

arXiv:2605.08321v1 Announce Type: cross Abstract: LLMs are increasingly capable of persuasion, which raises the question of how to protect users against manipulation. In a preregistered user study (N=

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Multi-Tier Labeling and Physics-Informed Learning for Orbital Anomaly Detection at Scale

DGX agent

arXiv:2605.09790v1 Announce Type: cross Abstract: Detecting orbital anomalies, such as maneuvers, atmospheric decay, and attitude upsets, across the rapidly growing population of low-Earth-orbit (LEO)

model-releasesarxiv-cs-ai
12 May 2026
Local Ai

On Distinguishing Capability Elicitation from Capability Creation in Post-Training: A Free-Energy Perspective

DGX agent

arXiv:2605.08368v1 Announce Type: new Abstract: Debates about large language model post-training often treat supervised fine-tuning (SFT) as imitation and reinforcement learning (RL) as discovery. But

local-aiarxiv-cs-ai
12 May 2026
Model Releases

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents

DGX agent

arXiv:2605.08876v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous agents that execute tool-augmented, multi-step tasks, where latency is a critical f

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Parameter-Efficient Neuroevolution for Diverse LLM Generation: Quality-Diversity Optimization via Prompt Embedding Evolution

DGX agent

arXiv:2605.09781v1 Announce Type: cross Abstract: Large Language Models exhibit mode collapse, producing homogeneous outputs that fail to explore valid solution spaces. We present QD-LLM, a framework

model-releasesarxiv-cs-ai
12 May 2026
Research

PLACO: A Multi-Stage Framework for Cost-Effective Performance in Human-AI Teams

DGX agent

arXiv:2605.08388v1 Announce Type: new Abstract: Human-AI teams play a pivotal role in improving overall system performance when neither the human nor the model can achieve such performance on their ow

researcharxiv-cs-ai
12 May 2026
Model Releases

PlantMarkerBench: A Multi-Species Benchmark for Evidence-Grounded Plant Marker Reasoning

DGX agent

arXiv:2605.10032v1 Announce Type: new Abstract: Cell-type-specific marker genes are fundamental to plant biology, yet existing resources primarily rely on curated databases or high-throughput studies

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

PrAg-PO: Prompt Augmented Policy Optimization for Robust and Diverse Mathematical Reasoning

DGX agent

arXiv:2602.03190v3 Announce Type: replace-cross Abstract: Reinforcement learning algorithms such as group-relative policy optimization (GRPO) have shown strong potential for improving the mathematical

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Preserving Foundational Capabilities in Flow-Matching VLAs through Conservative SFT

DGX agent

arXiv:2605.08879v1 Announce Type: new Abstract: Unconstrained fine-tuning of flow-matching Vision-Language-Action (VLA) models drives dense parameter overwrites, degrading pre-trained capabilities. We

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

Process Matters more than Output for Distinguishing Humans from Machines

DGX agent

arXiv:2605.06524v2 Announce Type: replace Abstract: Reliable human-machine discrimination is becoming increasingly important as large language models and autonomous agents are deployed in online setti

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Prompt-Activation Duality: Improving Activation Steering via Attention-Level Interventions

DGX agent

arXiv:2605.10664v1 Announce Type: cross Abstract: Activation steering controls language model behavior by adding directions to internal representations at inference time, but standard residual-stream

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…371372373374375…1074
Next →