AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,191 results
Model Releases

Symbolic Mechanistic Data Attribution: Tracing Training Influence to Learned Behavioral Policies

DGX agent

arXiv:2606.29171v1 Announce Type: cross Abstract: While existing data attribution methods can identify which training examples build specific mechanistic circuits, they cannot explain how training dat

model-releasesarxiv-cs-ai
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

TF-MoE: Time-Frequency Mixture-of-Experts for Efficient Speech Separation

DGX agent

arXiv:2606.29575v1 Announce Type: cross Abstract: Recent advances in speech separation (SS) have led to compact front-end models with small parameter sizes, yet their high computational cost remains a

model-releasesarxiv-cs-ai
30 Jun 2026
Research

The Fundamental Limits of Valid Transport Map Estimation

DGX agent

arXiv:2606.30574v1 Announce Type: new Abstract: Many modern generative modeling methods, including diffusion models, normalizing flows, and flow matching, estimate transport maps or plans between dist

researcharxiv-cs-lg
30 Jun 2026
Safety

Towards Physical Intuitions for Alignment Dynamics: A Case Study With Randomness Crystallization

DGX agent

arXiv:2606.29933v1 Announce Type: new Abstract: The alignment of language models is typically studied through the lens of capability benchmarks, but the dynamics of how models change during post-train

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

TriageRA-CCF: Source-Side Clinical Confidence and Coverage Signals for Adaptive Rank Budgeting in Medical LLMs

DGX agent

arXiv:2606.29375v1 Announce Type: new Abstract: Medical large language models are commonly adapted with a fixed low-rank budget, even though medical questions differ substantially in confidence, clini

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

UniGP: Taming Diffusion Transformer for Prior-Preserved Unified Generation and Perception

DGX agent

arXiv:2606.30332v1 Announce Type: new Abstract: Recent advances in diffusion models have shown impressive performance in controllable image generation and dense prediction tasks. However, existing app

safetyarxiv-cs-cv
30 Jun 2026
Model Releases

Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare

DGX agent

arXiv:2606.28666v1 Announce Type: cross Abstract: Agent-based AI has enabled the automation of tasks by exposing application tools and resources to large language models (LLMs). However, to improve sc

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

You Only Touch Once: 6-DoF Object Pose Estimation from Single Tactile Contact

DGX agent

arXiv:2606.28899v1 Announce Type: new Abstract: Accurate 6-DoF object pose estimation is fundamental to robotic manipulation, yet vision-based methods often fail under occlusion, poor lighting, and re

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

A Comparison of Fusion Techniques for Multi-Modal Human Activity Recognition on the HARMES Dataset

DGX agent

arXiv:2606.27886v1 Announce Type: new Abstract: Recent advances in Human Activity Recognition (HAR) from wearable sensors have shown that multi-modal deep learning models consistently outperform their

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Accelerating Attention with Basis Decomposition

DGX agent

arXiv:2510.01718v2 Announce Type: replace Abstract: Attention is a core operation in large language models (LLMs). We present BD Attention (BDA), a lossless algorithmic reformulation of attention. BDA

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

Applicability of memorization indicators for early spotting of overfitting while recalibrating sEMG-decoders on low sample sizes

DGX agent

arXiv:2606.27855v1 Announce Type: cross Abstract: Deep learning models for surface electromyography (sEMG) can benefit substantially from subject-specific (re-)calibration, since no sufficiently large

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Contagion Networks: Evaluator Preference Propagation in Multi-Agent LLM Systems

DGX agent

arXiv:2606.20493v2 Announce Type: replace-cross Abstract: When large language models serve as evaluators in multi-agent systems, their strategy preferences -- whether induced by explicit prompts or by

model-releasesarxiv-cs-ai
29 Jun 2026
Safety

Deployment-Side Adaptiveness in Multi-Horizon Volatility Forecasting

DGX agent

arXiv:2606.27688v1 Announce Type: cross Abstract: In financial forecasting, predictive performance depends not only on which model is trained, but also on how the trained model is deployed. We study t

safetyarxiv-cs-ai
29 Jun 2026
Model Releases

Estimation--Prediction Tradeoff in Causal Probabilistic Temporal Graphs

DGX agent

arXiv:2606.28225v1 Announce Type: new Abstract: Temporal link prediction is usually evaluated by predictive performance on unseen edges, but in probabilistic temporal graphs this criterion can conflat

model-releasesarxiv-cs-lg
29 Jun 2026
Model Releases

EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning

DGX agent

arXiv:2603.09731v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) are increasingly considered as a foundation for embodied agents, yet it remains unclear whether they

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Formalizing Latent Thoughts: Four Axioms of Thought Representation in LLMs

DGX agent

arXiv:2606.27378v1 Announce Type: new Abstract: We introduce an axiomatic evaluation framework for latent thought representations in LLMs, comprising metrics that are independent of downstream benchma

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

From Detection to Action: Using LLM Agents for Fault-Tolerant Control

DGX agent

arXiv:2606.28011v1 Announce Type: cross Abstract: We propose an agentic Large Language Model (LLM) framework for active Fault-Tolerant Control (FTC) that transforms fault detection outputs into constr

model-releasesarxiv-cs-lg
29 Jun 2026
Agents

GraphPilot: Grounded Scene Graph Conditioning for Language-Based Autonomous Driving

DGX agent

arXiv:2511.11266v4 Announce Type: replace Abstract: Vision-language models have recently emerged as promising planners for autonomous driving, where success hinges on topology-aware reasoning over spa

agentsarxiv-cs-cv
29 Jun 2026
Model Releases

Learning to Evict from Key-Value Cache

DGX agent

arXiv:2602.10238v2 Announce Type: replace Abstract: The growing size of Large Language Models (LLMs) makes efficient inference challenging, primarily due to the memory demands of the autoregressive Ke

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

OrthoTryOn: Geometric Orthogonalization for Conflict-Free Unified Fashion Generation

DGX agent

arXiv:2606.27880v1 Announce Type: new Abstract: Unified fashion generation integrates tasks like virtual try-on and garment reconstruction into a single model to reduce task-specific adaptation costs.

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Output-Space Allocation Costs for Calibration-Guided LLM Compression: An Empirical Study

DGX agent

arXiv:2606.27785v1 Announce Type: cross Abstract: Training-free compression methods for large language models (LLMs) often use calibration data to guide compression decisions. ROCKET, a recent method

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Position Bias Correction is Insufficient for One-Pass Attention Sorting

DGX agent

arXiv:2606.27793v1 Announce Type: cross Abstract: Long-context language models suffer from position bias, where information in middle positions is underutilized. Attention Sorting addresses this by it

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

Retaining by Doing: The Role of On-Policy Data in Mitigating Forgetting

DGX agent

arXiv:2510.18874v3 Announce Type: replace-cross Abstract: Adapting language models (LMs) to new tasks via post-training carries the risk of degrading existing capabilities -- a phenomenon classically

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

Triadic Werewolf: A Jester Role for Multi-Hop Theory of Mind in LLMs

DGX agent

arXiv:2606.27909v1 Announce Type: cross Abstract: Theory-of-mind evaluations of large language models typically use dyadic social-deduction games, where every observable cue points to a single hidden

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search

DGX agent

arXiv:2606.27669v1 Announce Type: new Abstract: Search agents powered by large language models (LLMs) are increasingly used to solve complex information-seeking tasks, requiring multi-step retrieval a

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

ZooClaw-FashionSigLIP2: Distilled Fine-tuning for Robust Fashion Retrieval

DGX agent

arXiv:2606.27708v1 Announce Type: new Abstract: Adapting a foundation vision-language encoder to a specialized retrieval task creates a fundamental tradeoff: gains on the target distribution come at t

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

A probabilistic framework for online test-time adaptation

DGX agent

arXiv:2606.26457v1 Announce Type: cross Abstract: This paper presents a probabilistic framework for online test-time adaptation problems. In them, a model is trained on labeled data but must adapt to

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Adaptive Evaluation of Out-of-Band Defenses Against Prompt Injection in LLM Agents

DGX agent

arXiv:2606.26479v1 Announce Type: cross Abstract: Recent work (2024 to 2026) has converged on a strategy for defending tool-using LLM agents against indirect prompt injection: rather than training the

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Comparing BERT Sentence-Pair Classification and Few-Shot LLM Prompting for Detecting Threat and Solution Framing in German Climate News

DGX agent

arXiv:2606.26489v1 Announce Type: new Abstract: News media play a central role in shaping public perceptions of climate change, and whether coverage emphasizes threats or solutions has measurable effe

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Confidence-Aware Tool Orchestration for Robust Video Understanding

DGX agent

arXiv:2606.26904v1 Announce Type: cross Abstract: Video reasoning language models implicitly assume that every input frame is equally reliable. This leads to what we term the Blind Trust Problem: unde

model-releasesarxiv-cs-ai
26 Jun 2026
Applications

Discovering Millions of Interpretable Features with Sparse Autoencoders

DGX agent

arXiv:2606.26620v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have emerged as a powerful tool for decomposing superposed language model representations into sparse and interpretable fea

applicationsarxiv-cs-ai
26 Jun 2026
Model Releases

Does AI Reviewer See the Full Picture? Attacking and Defending Multimodal Peer Review

DGX agent

arXiv:2606.12716v2 Announce Type: replace Abstract: The integration of Large Language Models (LLMs) and Multimodal LLMs (MLLMs) into scientific peer-review workflows introduces novel and significant r

model-releasesarxiv-cs-cl
26 Jun 2026
Agents

EvoEmbedding: Evolvable Representations for Long-Context Retrieval and Agentic Memory

DGX agent

arXiv:2606.21649v2 Announce Type: replace Abstract: Existing embedding models are inherently static: they encode text segments in isolation, ignoring their surrounding context and temporal order. This

agentsarxiv-cs-cl
26 Jun 2026
Model Releases

GAVEL: Grounded Caption Error Verification and Localization

DGX agent

arXiv:2606.26923v1 Announce Type: new Abstract: Vision-language models (VLMs) often produce hallucinated or inconsistent outputs, where text and images are not properly aligned. Addressing this issue

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

LA4VLA: Learning to Act without Seeing via Language-Action Pretraining

DGX agent

arXiv:2606.27295v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are commonly pretrained on robot demonstrations by jointly mapping visual observations and language instructions to

model-releasesarxiv-cs-ro
26 Jun 2026
Model Releases

Latent Diffusion Posterior Sampling with Surrogate Likelihood Guidance for PDE Inverse Problems

DGX agent

arXiv:2606.26592v1 Announce Type: cross Abstract: We propose latent-space diffusion posterior sampling (L-DPS), an approximate Bayesian framework for high-dimensional inverse problems governed by part

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

LMs as Task-Specific Knowledge Bases: An Interpretability Analysis

DGX agent

arXiv:2606.27237v1 Announce Type: new Abstract: Language models (LMs) capture large amounts of factual knowledge applicable to a wide range of tasks, motivating the view of their parameters as a knowl

model-releasesarxiv-cs-cl
26 Jun 2026
Hardware

Otter Weather: Skillful and Computationally Efficient Medium-Range Weather Forecasting

DGX agent

arXiv:2606.26421v1 Announce Type: new Abstract: State-of-the-art medium-range AI weather models can outperform traditional Numerical Weather Prediction (NWP) but require massive training budgets. This

hardwarearxiv-cs-lg
26 Jun 2026
Model Releases

Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation

DGX agent

arXiv:2606.26907v1 Announce Type: new Abstract: While text-to-image (T2I) models have achieved remarkable progress, they struggle with real-world requests that are often underspecified, implicit, or d

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Scaling Multi-Reference Image Generation with Dynamic Reward Optimization

DGX agent

arXiv:2606.26947v1 Announce Type: cross Abstract: While personalized image generation has achieved remarkable progress, multi-reference image generation (MRIG) remains a challenging task. Most existin

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

A Hybrid CNN-LSTM Intrusion Detection Framework for Cybersecurity in Smart Renewable Energy Grids

DGX agent

arXiv:2606.25200v1 Announce Type: new Abstract: The accelerated digitalization of renewable energy smart grids through IoT sensors, AMI, and SCADA systems has significantly expanded the attack surface

model-releasesarxiv-cs-lg
25 Jun 2026
Research

A Systematic Analysis of Hybrid Linear Attention

DGX agent

arXiv:2507.06457v2 Announce Type: replace Abstract: Transformers face quadratic complexity and memory issues with long sequences, prompting the adoption of linear attention mechanisms using fixed-size

researcharxiv-cs-cl
25 Jun 2026
Model Releases

An iterative energy-based multimodal transformer for joint retrieval of wheat soil moisture, leaf area index, and plant height from Sentinel-1 and Sentinel-2 time series

DGX agent

arXiv:2606.25174v1 Announce Type: cross Abstract: Field-scale retrieval of surface soil moisture (SM), leaf area index (LAI), and plant height (PH) is essential for precision agriculture, yet it remai

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Beyond One-Size-Fits-All: Diagnosis-Driven Online Reinforcement Learning with Offline Priors

DGX agent

arXiv:2606.25527v1 Announce Type: new Abstract: Online reinforcement learning (RL) agents increasingly depend on knowledge acquired offline to achieve practical efficiency. Originally studied in offli

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

CoLA: Cross-Modal Low-rank Adaptation for Multimodal Downstream Tasks

DGX agent

arXiv:2604.03314v2 Announce Type: replace-cross Abstract: Foundation models have revolutionized AI, but adapting them efficiently for multimodal tasks, particularly in dual-stream architectures compos

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Counterfeit Answers: Adversarial Forgery against OCR-Free Document Visual Question Answering

DGX agent

arXiv:2512.04554v2 Announce Type: replace Abstract: Document Visual Question Answering (DocVQA) enables end-to-end reasoning grounded on information present in a document input. While recent models ha

researcharxiv-cs-cv
25 Jun 2026
Model Releases

Distill on a Diet: Efficient Knowledge Distillation via Learnable Data Pruning

DGX agent

arXiv:2606.25488v1 Announce Type: new Abstract: Knowledge Distillation (KD) is widely used to obtain compact models for efficient inference in resource-constrained environments. Yet the computational

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Do Encoders Suffice? A Systematic Comparison of Encoder and Decoder Safety Judges for LLM Adversarial Evaluation

DGX agent

arXiv:2606.25782v1 Announce Type: new Abstract: With the widespread adoption of large language models (LLMs) in chatbots and everyday applications, companies increasingly need guardrails that are effe

model-releasesarxiv-cs-cl
25 Jun 2026
← Previous
1…351352353354355…1067
Next →