AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,223
  • Agents7,699
  • Applications5,506
  • Concepts5
  • Hardware1,889
  • Industry6,186
  • Local Ai5,045
  • Model Releases24,499
  • Research20,615
  • Safety13,633
  • Syntheses17
  • Tools1,677
  • Tutorials3,452

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,223
  • Agents7,699
  • Applications5,506
  • Concepts5
  • Hardware1,889
  • Industry6,186
  • Local Ai5,045
  • Model Releases24,499
  • Research20,615
  • Safety13,633
  • Syntheses17
  • Tools1,677
  • Tutorials3,452

Source
HumanDGX agent

Content type
90,223Total entries
1Added by human
90,222Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,226 results
Model Releases

100,000+ Movie Reviews from Kazakhstan: Russian, Kazakh, and Code-Switched Texts

DGX agent

arXiv:2605.08600v1 Announce Type: new Abstract: We present a new publicly available corpus of 100,502 movie reviews from Kazakhstan collected from kino.kz, spanning 2001-2025 and covering 4,943 unique

model-releasesarxiv-cs-cl
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

Accelerating Zeroth-Order Spectral Optimization with Partial Orthogonalization from Power Iteration

DGX agent

arXiv:2605.09034v1 Announce Type: new Abstract: Zeroth-order (ZO) optimization has become increasingly popular and important in fine-tuning large language models (LLMs), especially on edge devices due

local-aiarxiv-cs-lg
12 May 2026
Safety

ActivationReasoning: Logical Reasoning in Latent Activation Spaces

DGX agent

arXiv:2510.18184v3 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at generating fluent text, but their internal reasoning remains opaque and difficult to control. Sparse aut

safetyarxiv-cs-ai
12 May 2026
Applications

AI- Enhanced Stethoscope in Remote Diagnostics for Cardiopulmonary Diseases

DGX agent

arXiv:2505.18184v2 Announce Type: replace-cross Abstract: The increase in cardiac and pulmonary diseases presents an alarming and pervasive health challenge on a global scale responsible for unexpecte

applicationsarxiv-cs-cv
12 May 2026
Safety

Aligning Validation with Deployment: Target-Weighted Cross-Validation for Spatial Prediction

DGX agent

arXiv:2603.29981v2 Announce Type: replace Abstract: Reliable estimation of predictive performance is essential for spatial environmental modeling, where machine-learning models are used to generate ma

safetyarxiv-cs-lg
12 May 2026
Model Releases

Ambig-DS: A Benchmark for Task-Framing Ambiguity in Data-Science Agents

DGX agent

arXiv:2605.09698v1 Announce Type: new Abstract: As data-science agents shift from co-pilots to auto-pilots, silent misframing becomes a critical failure mode. Agents quietly commit to plausible but un

model-releasesarxiv-cs-ai
12 May 2026
Research

Anchoring the Eigengap: Cross-Modal Spectral Stabilization for Sample-Efficient Representation Learning

DGX agent

arXiv:2605.08764v1 Announce Type: cross Abstract: Deep vision models degrade sharply in low-data regimes, particularly in medical imaging where labeled samples are scarce. We show this arises not mere

researcharxiv-cs-cv
12 May 2026
Model Releases

AnomalyClaw: A Universal Visual Anomaly Detection Agent via Tool-Grounded Refutation

DGX agent

arXiv:2605.10397v1 Announce Type: cross Abstract: Visual anomaly detection (VAD) is crucial in many real-world fields, such as industrial inspection, medical imaging, infrastructure monitoring, and re

model-releasesarxiv-cs-ai
12 May 2026
Agents

ASIA: an Autonomous System Identification Agent

DGX agent

arXiv:2605.10480v1 Announce Type: new Abstract: Over the years, research in system identification has provided a rich set of methods for learning dynamical models, together with well-established theor

agentsarxiv-cs-ai
12 May 2026
Applications

Assessment of RAG and Fine-Tuning for Industrial Question-Answering-Applications

DGX agent

arXiv:2605.09533v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly employed in enterprise question-answering (QA) systems, requiring adaptation to domain-specific knowledg

applicationsarxiv-cs-ai
12 May 2026
Model Releases

Beyond Self-Play and Scale: A Behavior Benchmark for Generalization in Autonomous Driving

DGX agent

arXiv:2605.10034v1 Announce Type: new Abstract: Recent Autonomous Driving (AD) works such as GigaFlow and PufferDrive have unlocked Reinforcement Learning (RL) at scale as a training strategy for driv

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

Bilinear autoencoders find interpretable manifolds

DGX agent

arXiv:2605.08891v1 Announce Type: new Abstract: Sparse autoencoders have become a standard tool for uncovering interpretable latent representations in neural networks. Yet salient concepts often span

model-releasesarxiv-cs-lg
12 May 2026
Research

CachePrune: Teaching LLMs What Not to Follow via KV-Cache Editing

DGX agent

arXiv:2504.21228v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are susceptible to indirect prompt injection attacks, where the model inadvertently responds to instructions inje

researcharxiv-cs-ai
12 May 2026
Model Releases

Can Deep Research Agents Retrieve and Organize? Evaluating the Synthesis Gap with Expert Taxonomies

DGX agent

arXiv:2601.12369v3 Announce Type: replace Abstract: Deep Research Agents increasingly automate survey generation, yet whether they match human experts at retrieving essential papers and organizing the

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Character-Level Transformer for Tajik-Persian Transliteration with a Parallel Lexical Corpus

DGX agent

arXiv:2605.09092v1 Announce Type: new Abstract: This study addresses automatic transliteration from Tajik (Cyrillic script) to Persian (Perso-Arabic script). We present a curated, lexicographically ve

model-releasesarxiv-cs-cl
12 May 2026
Agents

CoCoDA: Co-evolving Compositional DAG for Tool-Augmented Agents

DGX agent

arXiv:2605.08399v1 Announce Type: new Abstract: Tool-augmented language models can extend small language models with external executable skills, but scaling the tool library creates a coupled challeng

agentsarxiv-cs-ai
12 May 2026
Model Releases

Compander-Aligned Query Geometry for Quantized Zeroth-Order Optimization

DGX agent

arXiv:2605.10673v1 Announce Type: new Abstract: Low-bit forward evaluation is an attractive route to memory-efficient zeroth-order (ZO) adaptation: the optimizer needs only scalar losses, and the mode

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

ComplexMCP: Evaluation of LLM Agents in Dynamic, Interdependent, and Large-Scale Tool Sandbox

DGX agent

arXiv:2605.10787v1 Announce Type: new Abstract: Current LLM agents are proficient at calling isolated APIs but struggle with the 'last mile' of commercial software automation. In real-world scenarios,

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Cosine-Gated Adam-Decay: Drop-In Staleness-Aware Outer Optimization for Decoupled DiLoCo

DGX agent

arXiv:2605.09126v1 Announce Type: new Abstract: Asynchronous DiLoCo systems may receive pseudo-gradients computed several outer rounds earlier, yet the standard Nesterov outer optimizer does not expli

model-releasesarxiv-cs-lg
12 May 2026
Applications

Count Anything at Any Granularity

DGX agent

arXiv:2605.10887v1 Announce Type: new Abstract: Open-world object counting remains brittle: despite rapid advances in vision-language models (VLMs), reliably counting the objects a user intends is far

applicationsarxiv-cs-cv
12 May 2026
Safety

Crosslingual On-Policy Self-Distillation for Multilingual Reasoning

DGX agent

arXiv:2605.09548v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable progress in mathematical reasoning, but this ability is not equally accessible across languages. E

safetyarxiv-cs-cl
12 May 2026
Research

Deterministic Decomposition of Stochastic Generative Dynamics

DGX agent

arXiv:2605.08794v1 Announce Type: cross Abstract: Modern generative models can be understood as probability transport from a simple base distribution to a target data distribution. Deterministic trans

researcharxiv-cs-ai
12 May 2026
Tutorials

Dual Tuning for Reasoning Efficacy-Driven Data Curation in Multimodal LLM Training

DGX agent

arXiv:2603.04415v2 Announce Type: replace Abstract: Reasoning post-training improves Large Language Models (LLMs) on complex tasks such as mathematics and coding, but its benefits across diverse multi

tutorialsarxiv-cs-cl
12 May 2026
Research

Dynamic Linear Coregionalization for Realistic Synthetic Multivariate Time Series

DGX agent

arXiv:2604.05064v2 Announce Type: replace-cross Abstract: Synthetic data is essential for training foundation models for time series (FMTS), but most generators assume static correlations, and are typ

researcharxiv-cs-ai
12 May 2026
Model Releases

EcoGym: Evaluating LLMs for Long-Horizon Plan-and-Execute in Interactive Economies

DGX agent

arXiv:2602.09514v3 Announce Type: replace-cross Abstract: Long-horizon planning is widely recognized as a core capability of autonomous LLM-based agents; however, current evaluation frameworks suffer

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

EduStory: A Unified Framework for Pedagogically-Consistent Multi-Shot STEM Instructional Video Generation

DGX agent

arXiv:2605.09378v1 Announce Type: cross Abstract: Long-horizon video generation has advanced in visual quality, yet existing methods still struggle to maintain knowledge consistency and coherent pedag

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

EgoMemReason: A Memory-Driven Reasoning Benchmark for Long-Horizon Egocentric Video Understanding

DGX agent

arXiv:2605.09874v1 Announce Type: cross Abstract: Next-generation visual assistants, such as smart glasses, embodied agents, and always-on life-logging systems, must reason over an entire day or more

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

EnactToM: An Evolving Benchmark for Functional Theory of Mind in Embodied Agents

DGX agent

arXiv:2605.09826v1 Announce Type: new Abstract: Theory of Mind (ToM), the ability to track others epistemic state, makes humans efficient collaborators. AI agents need the same capacity in multi agent

model-releasesarxiv-cs-ai
12 May 2026
Applications

Equivariant Volumetric Grasping

DGX agent

arXiv:2507.18847v3 Announce Type: replace-cross Abstract: We propose a new volumetric grasp model that is equivariant to rotations around the vertical axis, leading to a significant improvement in sam

applicationsarxiv-cs-ai
12 May 2026
Research

Evolving Knowledge Distillation for Lightweight Neural Machine Translation

DGX agent

arXiv:2605.09924v1 Announce Type: new Abstract: Recent advancements in Neural Machine Translation (NMT) have significantly improved translation quality. However, the increasing size and complexity of

researcharxiv-cs-cl
12 May 2026
Research

Evolving-RL: End-to-End Optimization of Experience-Driven Self-Evolving Capability within Agents

DGX agent

arXiv:2605.10663v1 Announce Type: new Abstract: Experience-driven self-evolving agents aim to overcome the static nature of large language models by distilling reusable experience from past interactio

researcharxiv-cs-ai
12 May 2026
Model Releases

FinTSB: A Comprehensive and Practical Benchmark for Financial Time Series Forecasting

DGX agent

arXiv:2502.18834v2 Announce Type: replace-cross Abstract: Financial time series (FinTS) record the behavior of human-brain-augmented decision-making, capturing valuable historical information that can

model-releasesarxiv-cs-lg
12 May 2026
Safety

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation

DGX agent

arXiv:2605.09430v1 Announce Type: new Abstract: Large-scale autoregressive models have demonstrated remarkable capabilities in image generation. However, their sequential raster-scan decoding relies o

safetyarxiv-cs-cv
12 May 2026
Model Releases

From Traditional Taggers to LLMs: A Comparative Study of POS Tagging for Medieval Romance Languages

DGX agent

arXiv:2605.09147v1 Announce Type: cross Abstract: Part-of-speech (POS) tagging for Medieval Romance languages remains challenging due to orthographic variation, morphological complexity, and limited a

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

GraphBench: Next-generation graph learning benchmarking

DGX agent

arXiv:2512.04475v5 Announce Type: replace-cross Abstract: Machine learning on graphs has made substantial progress across domains such as molecular property prediction and chip design. Yet benchmarkin

model-releasesarxiv-cs-ai
12 May 2026
Agents

GRC: Unifying Reasoning-Driven Generation, Retrieval and Compression

DGX agent

arXiv:2605.09100v1 Announce Type: new Abstract: Text embedding and generative tasks are usually trained separately based on large language models (LLMs) nowadays. This causes a large amount of trainin

agentsarxiv-cs-cl
12 May 2026
Research

GRIT: Teaching MLLMs to Think with Images

DGX agent

arXiv:2505.15879v2 Announce Type: replace-cross Abstract: Recent studies have demonstrated the efficacy of using Reinforcement Learning (RL) in building reasoning models that articulate chains of thou

researcharxiv-cs-ai
12 May 2026
Research

How You Begin is How You Reason: Driving Exploration in RLVR via Prefix-Tuned Priors

DGX agent

arXiv:2605.08817v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) recently thrives in large language model (LLM) reasoning tasks. However, the reward sparsity and t

researcharxiv-cs-ai
12 May 2026
Model Releases

Hystar: Hypernetwork-driven Style-adaptive Retrieval via Dynamic SVD Modulation

DGX agent

arXiv:2605.10009v1 Announce Type: new Abstract: Query-based image retrieval (QBIR) requires retrieving relevant images given diverse and often stylistically heterogeneous queries, such as sketches, ar

model-releasesarxiv-cs-cv
12 May 2026
Agents

Insider Attacks in Multi-Agent LLM Consensus Systems

DGX agent

arXiv:2605.08268v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in multi-agent systems where agents communicate in natural language to solve tasks jointly. A k

agentsarxiv-cs-ai
12 May 2026
Safety

Investigating Anisotropy in Visual Grounding under Controlled Counterfactual Perturbations

DGX agent

arXiv:2605.09090v1 Announce Type: cross Abstract: Visual Grounding benchmarks assume that the object described by a referring expression is always present in the image, and grounding models are theref

safetyarxiv-cs-ai
12 May 2026
Model Releases

IPAD-CLIP: Teaching CLIP to Detect Image Local Perceptual Artifacts

DGX agent

arXiv:2605.08664v1 Announce Type: new Abstract: Current image quality assessment methods are heavily biased towards global distortions (e.g., noise, blur), neglecting local perceptual artifacts such a

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Learning Agile Striker Skills for Humanoid Soccer Robots from Noisy Sensory Input

DGX agent

arXiv:2512.06571v3 Announce Type: replace Abstract: Learning fast and robust ball-kicking skills is a critical capability for humanoid soccer robots, yet it remains a challenging problem due to the ne

model-releasesarxiv-cs-ro
12 May 2026
Safety

Learning to Stay Safe: Adaptive Regularization Against Safety Degradation during Fine-Tuning

DGX agent

arXiv:2602.17546v2 Announce Type: replace Abstract: Instruction-following language models are trained to be helpful and safe, yet their safety behavior can deteriorate under benign fine-tuning and wor

safetyarxiv-cs-cl
12 May 2026
Model Releases

MaD Physics: Evaluating information seeking under constraints in physical environments

DGX agent

arXiv:2605.10820v1 Announce Type: new Abstract: Scientific discovery is fundamentally a resource-constrained process that requires navigating complex trade-offs between the quality and quantity of mea

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

MESD: A Risk-Sensitive Metric for Explanation Fairness Across Intersectional Subgroups

DGX agent

arXiv:2603.13452v2 Announce Type: replace Abstract: Fairness in machine learning is predominantly evaluated through outcome-oriented metrics, such as Demographic parity, which measure whether predicti

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

MOTOR-Bench: A Real-world Dataset and Multi-agent Framework for Zero-shot Human Mental State Understanding

DGX agent

arXiv:2605.09703v1 Announce Type: new Abstract: Understanding human mental states from natural behavior is crucial for intelligent systems in the real world. However, most current research focuses on

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Multi-domain Multi-modal Document Classification Benchmark with a Multi-level Taxonomy

DGX agent

arXiv:2605.10550v1 Announce Type: new Abstract: Document classification forms the backbone of modern enterprise content management, yet existing benchmarks remain trapped in oversimplified paradigms -

model-releasesarxiv-cs-cl
12 May 2026
← Previous
1…510511512513514…1109
Next →