AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,935 results
Research

GateKD: Confidence-Gated Closed-Loop Distillation for Robust Reasoning

DGX agent

arXiv:2605.13136v2 Announce Type: replace Abstract: Distilling multi-step reasoning abilities from large language models (LLMs) into compact student models remains challenging due to noisy rationales,

researcharxiv-cs-cl
2 Jun 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Global PIQA: Evaluating Commonsense Reasoning Across 100+ Languages and Cultures

DGX agent

arXiv:2510.24081v2 Announce Type: replace Abstract: To date, there exist almost no culturally-specific evaluation benchmarks for large language models (LLMs) that cover a large number of languages and

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Hierarchical Online Prompt Mutation with Dual-Loop Feedback for Guardrailed Evidence Document Generation: A Production-Evaluation Case Study

DGX agent

arXiv:2606.01472v1 Announce Type: cross Abstract: High-stakes production document-generation systems require language models to be adaptive, evidence-grounded, and auditable. We present HOPM, a hierar

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

HomeFlow: A Data Flywheel for Smart Home Agent Training with Verifiable Simulation

DGX agent

arXiv:2606.01230v1 Announce Type: new Abstract: Large language model agents are moving beyond text-only interaction toward physical-world control, with smart homes as a representative domain. Real dom

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Hybrid Imbalanced Regression Through Unified Data-Level and Algorithm-Level Balancing

DGX agent

arXiv:2606.01221v1 Announce Type: cross Abstract: Imbalanced learning is a critical challenge in machine learning, where underrepresented target values can bias models and degrade prediction performan

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Ideas in Inference-time Scaling can Benefit Generative Pre-training Algorithms

DGX agent

arXiv:2503.07154v3 Announce Type: replace-cross Abstract: Generative pre-training is often framed through a false dichotomy between autoregressive models for discrete signals and diffusion models for

researcharxiv-cs-ai
2 Jun 2026
Model Releases

InsightVQA: High-Dimensional Emotion-Cognitive Visual Question Answering Benchmark

DGX agent

arXiv:2606.02171v1 Announce Type: new Abstract: Visual emotion understanding requires models not only to recognize emotional states, but also to why they arise and perform higher-level cognitive reaso

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

LALE: Lightweight-Transformer Architecture for Land-Cover Estimation

DGX agent

arXiv:2606.02092v1 Announce Type: cross Abstract: Semantic segmentation of remote sensing imagery requires models that capture both global context and local detail under tight computational budgets. P

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

LaSR: Context-Aware Speech Recognition via Latent Reasoning

DGX agent

arXiv:2606.00507v1 Announce Type: new Abstract: Recent advances in Speech Large Language Models (Speech LLMs) have significantly enhanced spoken language understanding and reasoning. However, their co

model-releasesarxiv-cs-cl
2 Jun 2026
Agents

Latent Collaboration in Multi-Agent Systems

DGX agent

arXiv:2511.20639v3 Announce Type: replace-cross Abstract: Multi-agent systems (MAS) extend large language models (LLMs) from independent single-model reasoning to coordinative system-level intelligenc

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

LeAP: Learnable Adaptive Permutation for Feature Selection in Heterogeneous and Sparse Recommender Systems

DGX agent

arXiv:2606.01111v1 Announce Type: new Abstract: Modern industrial recommender systems rely on thousands of heterogeneous features -- ranging from low-dimensional scalars (e.g., statistical value) to h

model-releasesarxiv-cs-lg
2 Jun 2026
Research

Learning from Saturated Data: Signals Beyond Correctness for LLM Training

DGX agent

arXiv:2606.01436v1 Announce Type: new Abstract: The growing capabilities of large language models (LLMs) have led to the saturation of many benchmarks and training datasets used to improve them. Motiv

researcharxiv-cs-cl
2 Jun 2026
Model Releases

LLM Consortium for Software Design Refinement: A Controlled Experiment on Multi-Agent Collaboration Topologies

DGX agent

arXiv:2606.01490v1 Announce Type: cross Abstract: We present a controlled experiment evaluating 12 multi-agent LLM collaboration topologies for software architecture design. Using a 2imes2imes2 factor

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

LLM4Cov: Execution-Aware Agentic Learning for High-coverage Testbench Generation

DGX agent

arXiv:2602.16953v3 Announce Type: replace Abstract: Execution-aware LLM agents offer a promising paradigm for learning from tool feedback, but such feedback can be expensive and slow to obtain, making

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

LLMs for Cardiovascular Risk Prediction from Structured Clinical Data

DGX agent

arXiv:2606.00031v1 Announce Type: cross Abstract: Coronary artery disease (CAD) remains one of the leading causes of death globally, highlighting the need for reliable predictive systems to support ea

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Make a Video Call with LLM: A Measurement Campaign over Six Mainstream Apps

DGX agent

arXiv:2510.00481v2 Announce Type: replace-cross Abstract: In 2025, Large Language Model (LLM) services have launched a new feature -- AI video chat -- allowing users to interact with AI agents via rea

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Med-Scout: Curing MLLMs' Geometric Blindness in Medical Perception via Geometry-Aware RL Post-Training

DGX agent

arXiv:2601.23220v2 Announce Type: replace-cross Abstract: Despite recent Multimodal Large Language Models (MLLMs)' linguistic prowess in medical diagnosis, we find even state-of-the-art MLLMs suffer f

model-releasesarxiv-cs-ai
2 Jun 2026
Applications

MineDraft: A Framework for Batch Parallel Speculative Decoding

DGX agent

arXiv:2603.18016v2 Announce Type: replace-cross Abstract: Speculative decoding (SD) accelerates large language model inference by using a smaller draft model to propose draft tokens that are subsequen

applicationsarxiv-cs-ai
2 Jun 2026
Model Releases

MixerSENet: A Lightweight Framework for Efficient Hyperspectral Image Classification

DGX agent

arXiv:2606.01700v1 Announce Type: new Abstract: In this paper, a novel framework, MixerSENet, is introduced for hyperspectral image (HSI) classification, designed to address the challenges of computat

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

MMDG-Bench: A Benchmark for Multimodal Domain Generalization

DGX agent

arXiv:2606.00891v1 Announce Type: new Abstract: Multi-modal Domain Generalization (MMDG) seeks to leverage complementary modalities to enhance model robustness on unseen domains. Despite extensive pro

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

MOC: Multi-Order Communication in LLM-based Multi-Agent Systems

DGX agent

arXiv:2606.02359v1 Announce Type: new Abstract: Despite the remarkable progress of Large Language Model (LLM) based Multi-Agent Systems, most research focuses on optimizing coordination topology while

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

MomentKV: Closing the Directional Gap in KV Cache Eviction for Long-Context Inference

DGX agent

arXiv:2606.01563v1 Announce Type: new Abstract: Autoregressive decoding in Transformer-based language models relies on the KV cache, whose memory footprint grows linearly with sequence length and beco

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Move the Query, Not the Cache: Characterizing Cross-Instance Latent Attention Redistribution Across GPU Fabrics

DGX agent

arXiv:2606.01502v1 Announce Type: cross Abstract: Frontier LLMs increasingly decide what a query attends to with a sparse-attention indexer that picks a few KV-cache blocks per query: attention's unit

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Multimodal Action Diffusion for Robust End-to-End Autonomous Driving

DGX agent

arXiv:2606.02105v1 Announce Type: new Abstract: End-to-End Autonomous Driving (E2E-AD) systems have largely converged on predicting intermediate trajectory waypoints, delegating final control to hand-

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Multimodal Approaches for Visually-Rich Document Type Classification: A Comparative Analysis

DGX agent

arXiv:2606.02162v1 Announce Type: cross Abstract: Document type classification in visually rich documents remains challenging, as relevant information is distributed across textual, visual, and layout

model-releasesarxiv-cs-ai
2 Jun 2026
Local Ai

Multimodal Function Vectors for Visual Relations

DGX agent

arXiv:2510.02528v2 Announce Type: replace Abstract: Large Multimodal Models (LMMs) demonstrate impressive in-context learning abilities from few multimodal demonstrations, yet the internal mechanisms

local-aiarxiv-cs-ai
2 Jun 2026
Research

Not All Explanations Simulate Equally: Comparing Verbalized Feature Attributions and Self-Generated Rationales

DGX agent

arXiv:2606.01148v1 Announce Type: new Abstract: Natural-language explanations are often treated as a unified interface for understanding model behavior, but different explanation sources may support s

researcharxiv-cs-cl
2 Jun 2026
Model Releases

OneVLA: A Unified Framework for Embodied Tasks

DGX agent

arXiv:2606.01241v1 Announce Type: new Abstract: Navigation and manipulation are fundamental capabilities of embodied intelligence, enabling robots to interpret natural language commands and interact p

model-releasesarxiv-cs-ro
2 Jun 2026
Safety

OPD+: Rethinking the Advantage Design for On-Policy Distillation

DGX agent

arXiv:2606.01039v1 Announce Type: cross Abstract: On-policy distillation (OPD) is a widely used technique to transfer capabilities from capable teacher language models to the base student models, and

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Parameter-efficient Dual-encoder Architecture with Differentiable Choquet Integral Fusion for Underwater Acoustic Classification

DGX agent

arXiv:2606.02341v1 Announce Type: cross Abstract: Underwater acoustic classification has a wide array of oceanic applications, but faces challenges due to an increasingly complex acoustic environment.

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Personalized 3D Myocardial Infarct Geometry Reconstruction from Cine MRI for Cardiac Digital Twins

DGX agent

arXiv:2606.01808v1 Announce Type: new Abstract: Accurate 3D geometric characterization of myocardial infarction (MI) is essential for building cardiac digital twins (CDTs) to precisely simulate infarc

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Physics from Video: Identifiability of Time-Invariant Second-Order ODEs under Minimal Trajectory Conditions

DGX agent

arXiv:2606.00115v1 Announce Type: new Abstract: Bridging the gap between visual realism and physical understanding is a core challenge for video-based world models. We study the structural identifiabi

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

POIROT: Interrogating Agents for Failure Detection in Multi-Agent Systems

DGX agent

arXiv:2606.02282v1 Announce Type: new Abstract: Orchestrating Large Language Models into Multi-Agent Systems (LLM-MAS) has unlocked remarkable reasoning capabilities, yet emergent failures and halluci

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Post-Deterministic Distributed Systems: A New Foundation for Trustworthy Autonomous Infrastructure

DGX agent

arXiv:2606.01722v1 Announce Type: cross Abstract: For decades, distributed systems have typically assumed that correct participants execute protocol-specified behavior with stable, externally defined,

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

ProtStructQA: A Denotation Threshold in Protein Structural Reasoning

DGX agent

arXiv:2606.00451v1 Announce Type: new Abstract: Protein-language systems are often evaluated by whether they generate plausible biological text, but a structural question has a sharper semantics: it d

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Quality-Diversity Evolution for Discovering Diverse Vulnerabilities in LLM Safety

DGX agent

arXiv:2606.00801v1 Announce Type: cross Abstract: Current approaches to LLM adversarial testing suffer from coverage gaps: manual red-teaming does not scale, LLM-as-attacker methods exhibit mode colla

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

ResNet-34 with Lightweight Decoder for Accurate and Efficient Segmentation of Fetal Brain MRI

DGX agent

arXiv:2606.01293v1 Announce Type: cross Abstract: Accurate segmentation of fetal brain tissues in Magnetic Resonance Imaging (MRI) is critical for early diagnosis of congenital abnormalities and impro

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Rethinking RL Evaluation: Can Benchmarks Truly Reveal Failures of RL Methods?

DGX agent

arXiv:2510.10541v2 Announce Type: replace-cross Abstract: Current benchmarks are inadequate for evaluating progress in reinforcement learning (RL) for large language models (LLMs).Despite recent bench

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

RoboStressBench: Benchmarking VLM Robustness to Physical Visual Stress in Embodied Scenes

DGX agent

arXiv:2606.00828v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown strong visual understanding and are increasingly deployed in embodied AI systems, where reliable perception und

model-releasesarxiv-cs-cv
2 Jun 2026
Tutorials

Robust Reasoning via Dynamic Token Selection for Distribution-Aligned Self-Distillation

DGX agent

arXiv:2606.00628v1 Announce Type: new Abstract: Self-distillation improves learning efficiency by rewriting reference answers as training data that better matches the model's own distribution. However

tutorialsarxiv-cs-cl
2 Jun 2026
Model Releases

ROGLE: Robust Global-Local Alignment with Automated Region Supervision for Text-Based Person Search

DGX agent

arXiv:2606.01825v1 Announce Type: new Abstract: Text-Based Person Search (TBPS) aims to retrieve pedestrian images using natural language queries. However, existing TBPS models, especially those based

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

ROGUE: Misaligned Agent Behavior Arising from Ordinary Computer Use

DGX agent

arXiv:2606.00341v1 Announce Type: cross Abstract: As AI agents are increasingly deployed in real personal and corporate settings (email accounts, development workflows, company databases, etc.), safet

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

RoleCDE:Benchmarking and Mitigating Role-Alignment Trade-offs in Role-Playing Agents

DGX agent

arXiv:2606.01552v1 Announce Type: new Abstract: Role-playing agents(RPAs) are widely used to steer large language models(LLMs) toward role-consistent behavior, yet existing benchmarks mainly evaluate

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

SafeMCP: Proactive Power Regulation for LLM Agent Defense via Environment-Grounded Look-Ahead Reasoning

DGX agent

arXiv:2606.01991v1 Announce Type: new Abstract: As Large Language Model (LLM) agents increasingly leverage the Model Context Protocol (MCP) to operate in complex environments, the expansion of their a

safetyarxiv-cs-ai
2 Jun 2026
Tutorials

SCL: Towards Domain Generalization via Single-Temporal Multimodal Contrastive Learning for Remote Sensing Change Detection

DGX agent

arXiv:2404.11326v5 Announce Type: replace Abstract: In recent years, change detection and anomaly detection models based on CNN and transformer have achieved remarkable success across various datasets

tutorialsarxiv-cs-cv
2 Jun 2026
Research

Score imes Decoder: A Unified View of Unsupervised Inference-Time Scaling for Hallucination Mitigation

DGX agent

arXiv:2606.00739v1 Announce Type: new Abstract: Large language models hallucinate even when the answer lies within their parameters. While inference-time scaling can surface this latent knowledge, the

researcharxiv-cs-lg
2 Jun 2026
Agents

Seq-DeepIPC: Sequential Sensing for End-to-End Control in Legged Robot Navigation

DGX agent

arXiv:2510.23057v2 Announce Type: replace-cross Abstract: We present Seq-DeepIPC, a sequential end-to-end perception-to-control model for legged robot navigation in real-world environments. Seq-DeepIP

agentsarxiv-cs-cv
2 Jun 2026
Model Releases

SHARP: Sleep-based Hierarchical Accelerated Replay for Long Range Non-Stationary Temporal Pattern Recognition

DGX agent

arXiv:2606.00732v1 Announce Type: new Abstract: Learning long-range non-stationary temporal patterns remains a core challenge for modern sequence models, particularly in strict streaming settings. In

model-releasesarxiv-cs-ai
2 Jun 2026
← Previous
1…419420421422423…1082
Next →