AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,499 results
22 May 2026

Why Semantic Entropy Fails: Geometry-Aware and Calibrated Uncertainty for Policy Optimization

SafetyDGX agent

arXiv:2605.21801v1 Announce Type: cross Abstract: Post-training has become central to improving reasoning and alignment in large language models, where critic-free models enable scalable learning from

X-Token: Projection-Guided Cross-Tokenizer Knowledge Distillation

Model ReleasesDGX agent

arXiv:2605.21699v1 Announce Type: cross Abstract: Cross-tokenizer knowledge distillation allows a student model to learn from teachers with incompatible vocabularies. Prior work operates on hidden sta

21 May 2026

Ada2MS: A Hybrid Optimization Algorithm Based on Exponential Mixing of Elementwise and Global Second-Moment Estimates

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2605.20533v1 Announce Type: new Abstract: Optimization algorithms are core methods by which machine learning models iteratively minimize loss functions, update parameters, learn from data, and i

AVSD: Adaptive-View Self-Distillation by Balancing Consensus and Teacher-Specific Privileged Signals

SafetyDGX agent

arXiv:2605.20643v1 Announce Type: cross Abstract: Self-distillation enables language models to learn on-policy from their own trajectories by using the same model as both student and teacher, with the

Bugcrowd launches reinforcement learning environments to train AI on real software vulnerabilities

Model ReleasesDGX agent

Crowdsourced cybersecurity company Bugcrowd Inc. today launched Reinforcement Learning Environments, a new offering that lets frontier artificial intelligence labs train models on real vulnerable soft

Build a Coding Assistant with Weaviate MCP: RAG over Code & Docs

Model ReleasesDGX agent

This guide demonstrates how to build a coding assistant using Weaviate's Model Context Protocol (MCP) integration, enabling retrieval-augmented generation (RAG) capabilities over codebases and documen

Causal Unlearning in Collaborative Optimization: Exact and Approximate Influence Reversal under Adversarial Contributions

Model ReleasesDGX agent

arXiv:2605.20341v1 Announce Type: new Abstract: Federated learning systems must support data deletion requests to comply with privacy regulations, yet retraining from scratch after each deletion is co

DIVE: Embedding Compression via Self-Limiting Gradient Updates

Model ReleasesDGX agent

arXiv:2605.20689v1 Announce Type: new Abstract: High-dimensional embeddings from large language models impose significant storage and computational costs on vector search systems. Recent embedding com

Divide et Calibra: Multiclass Local Calibration via Vector Quantization

Model ReleasesDGX agent

arXiv:2605.21060v1 Announce Type: new Abstract: Accurate and well-calibrated Machine Learning (ML) models are mandatory in high-stakes settings, yet effective multiclass calibration remains challengin

Evolutionary Generation of Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2602.06511v3 Announce Type: replace Abstract: Large language model (LLM)-based multi-agent systems (MAS) show strong promise for complex reasoning, planning, and tool-augmented tasks, but design

Federated LoRA Fine-Tuning for LLMs via Collaborative Alignment

Model ReleasesDGX agent

arXiv:2605.21217v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has emerged as a powerful tool for parameter-efficient fine-tuning of large language models (LLMs). This paper studies LoRA

Fine-grained Claim-level RAG Benchmark for Law

Model ReleasesDGX agent

arXiv:2605.21071v1 Announce Type: new Abstract: The rapid progress of large language models (LLMs) is shifting semantic search toward a question-answering paradigm, where users ask questions and LLMs

Geometry-Lite: Interpretable Safety Probing via Layer-Wise Margin Geometry

Model ReleasesDGX agent

arXiv:2605.20241v1 Announce Type: cross Abstract: Prompt-level safety probes for large language models use hidden-state representations to separate safe from unsafe prompts, but strong average detecti

In the next version of Claude Code: run /usage to see a breakdown of which Skills, Agents, MCPs, and Plugins are using your tokens CLI today…

Model ReleasesDGX agent

The next version of Claude Code will introduce a `/usage` command that provides a detailed breakdown of token consumption across different components including Skills, Agents, MCPs (Model Context Prot

LEAP: A closed-loop framework for perovskite precursor additive discovery

Model ReleasesDGX agent

arXiv:2605.20242v1 Announce Type: new Abstract: Efficient discovery of precursor additives is essential for improving the performance of perovskite solar cells, yet the large chemical space makes conv

Memory-Efficient Partitioned DNN Inference on Resource-Constrained Android Crowds

Model ReleasesDGX agent

arXiv:2605.20723v1 Announce Type: new Abstract: Deploying large deep neural networks on memory-constrained mobile devices is a central challenge in edge ML. While compression, pruning, and quantizatio

Open source 🤝 NVIDIA

Model ReleasesDGX agent

Open source 🤝 NVIDIA 👏 Congratulations to @cohere on Command A+ — a powerful new model optimized for NVIDIA Blackwell and trained using NVIDIA CUDA-X libraries. Proud to be a part of it! Learn more ⤵️

Preserve, Reveal, Expand: Faithful 4D Video Editing with Region-Aware Conditioning

Model ReleasesDGX agent

arXiv:2605.20961v1 Announce Type: new Abstract: Existing 4D-driven video diffusion models primarily target plausible generation, but faithful 4D editing requires preserving source-observed regions whi

Q-DiT4SR: Exploration of Detail-Preserving Diffusion Transformer Quantization for Real-World Image Super-Resolution

Model ReleasesDGX agent

arXiv:2602.01273v4 Announce Type: replace Abstract: Recently, Diffusion Transformers (DiTs) have emerged in Real-World Image Super-Resolution (Real-ISR) to generate high-quality textures, yet their he

Reasoning-Trace Collapse: Evaluating the Loss of Explicit Reasoning During Fine-Tuning

ResearchDGX agent

arXiv:2605.21127v1 Announce Type: new Abstract: Explicit reasoning models are trained to produce intermediate reasoning traces before final answers, but downstream fine-tuning is often performed on or

Sequential Data Augmentation for Generative Recommendation

Model ReleasesDGX agent

arXiv:2509.13648v3 Announce Type: replace Abstract: Generative recommendation plays a crucial role in personalized systems, predicting users' future interactions from their historical behavior sequenc

Verifiable Provenance and Watermarking for Generative AI: An Evidentiary Framework for International Operational Law and Domestic Courts

Model ReleasesDGX agent

arXiv:2605.21002v1 Announce Type: cross Abstract: Generative artificial intelligence now synthesizes photorealistic imagery, audio, and video at a cost that defeats traditional forensic intuition. The

VersusQ: Pairwise Margin Reasoning for Generalizable Video Quality Assessment

Model ReleasesDGX agent

arXiv:2605.21130v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have shown promise for video quality assessment, but most methods still predict an absolute score for each video. Such po

Vision Transformers and Convolutional Neural Networks for Land Use Scene Classification

Model ReleasesDGX agent

arXiv:2605.21268v1 Announce Type: new Abstract: Land Use Scene Classification (LUSC) from remote sensing imagery plays a critical role in environmental monitoring, urban planning, and sustainable reso

WaveGraphNet: Physics-Consistent Guided-Wave Damage Localization through Coupled Inverse-Forward Graph Learning

Model ReleasesDGX agent

arXiv:2605.20311v1 Announce Type: new Abstract: Guided-wave structural health monitoring enables damage localization in composite plates using sparse networks of bonded piezoelectric transducers. Howe

Weight Decay Regimes in Grokking Transformers: Cheap Online Diagnostics

Model ReleasesDGX agent

arXiv:2605.20441v1 Announce Type: new Abstract: Transformers trained on modular arithmetic exhibit sharp transitions between memorization, generalization, and collapse. We show that weight decay acts

WikiVQABench: A Knowledge-Grounded Visual Question Answering Benchmark from Wikipedia and Wikidata

Model ReleasesDGX agent

arXiv:2605.21479v1 Announce Type: new Abstract: Visual Question Answering (VQA) benchmarks have largely emphasized perception-based tasks that can be solved from visual content alone. In contrast, man

20 May 2026

100 things we announced at I/O 2026

Model ReleasesDGX agent

Google I/O 2026 unveiled new models, agents and tools to help users build, search, create, discover, shop and get more done. Key announcements included Gemini Omni, Google Antigravity, and Universal C

A Family of Divergence Measures for Evaluating the Reconstruction Quality of Explainable Ensemble Trees

Model ReleasesDGX agent

arXiv:2605.19618v1 Announce Type: new Abstract: Validating interpretable surrogate models for ensemble learners requires measuring agreement between the ensemble's internal representation and its surr

Acoustic scattering AI for non-invasive object classifications: A case study on hair assessment

Model ReleasesDGX agent

arXiv:2506.14148v2 Announce Type: replace-cross Abstract: This paper presents a novel non-invasive object classification approach using acoustic scattering, demonstrated through a case study on hair a

Adynamical systems view of training generativemodels and the memorization phenomenon

ResearchDGX agent

arXiv:2605.19483v1 Announce Type: new Abstract: Using recent works of one of the authors (VSB) on collapse in generative models and two time scale dynamics in stochastic gradient descent in high dimen

As always, 'hermes update'

ResearchDGX agent

Nous Research announced an update to their Hermes model, likely detailing improvements, new features, or performance enhancements to their open-source language model. The post was shared on X (formerl

Causal Evidence for Attention Head Imbalance in Modality Conflict Hallucination

Model ReleasesDGX agent

arXiv:2605.19250v1 Announce Type: new Abstract: Modality-conflict hallucination occurs when multimodal large language models (MLLMs) prioritize erroneous textual premises over contradictory visual evi

Compositional Literary Primitives in Instruction-Tuned LLMs: Cross-Architectural SAE Features for Self, Style, and Affect

Model ReleasesDGX agent

arXiv:2605.18808v1 Announce Type: cross Abstract: We characterize a compositional architecture of literary primitives in two instruction-tuned large language models (Llama 3.1 8B-Instruct and Gemma 2

DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows

Model ReleasesDGX agent

arXiv:2605.19099v1 Announce Type: new Abstract: We introduce DecisionBench, a benchmark substrate for emergent delegation in long-horizon agentic workflows. The substrate fixes a task suite (GAIA, tau

Depth2Pose: A Pose-Based Benchmark for Monocular Depth Estimation without Ground-Truth Depth

Model ReleasesDGX agent

arXiv:2605.19797v1 Announce Type: new Abstract: Monocular depth estimation has improved significantly in recent years, driven by increasingly powerful models and large-scale training data. Predicted d

Descriptive versus Regulatory Uncertainty in Bounded Predictive Systems

Local AiDGX agent

arXiv:2605.18909v1 Announce Type: new Abstract: Any system that models the world under finite representational capacity must compress; any compression entails a prior; and the prior is the system's bi

Differential-Integral Neural Operator for Long-Term Turbulence Forecasting

Model ReleasesDGX agent

arXiv:2509.21196v3 Announce Type: replace-cross Abstract: Accurately forecasting the long-term evolution of turbulence represents a grand challenge in scientific computing and is crucial for applicati

DMN: A Compositional Framework for Jailbreaking Multimodal LLMs with Multi-Image Inputs

Model ReleasesDGX agent

arXiv:2605.18915v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are vulnerable to jailbreak attacks, which can elicit harmful responses from MLLMs. Many MLLMs support multi-

DocQT: Improving Document Forgery Localization Robustness via Diverse JPEG Quantization Tables

Model ReleasesDGX agent

arXiv:2605.19688v1 Announce Type: new Abstract: Document manipulation localization models achieve strong performance on public benchmarks yet fail to generalize to operational document workflows. We i

EventPrune: Cascaded Event-Assisted Token Pruning for Efficient First-Person Dynamic Spatial Reasoning

Model ReleasesDGX agent

arXiv:2605.19506v1 Announce Type: new Abstract: First-person dynamic spatial reasoning requires models to track continuous motion and precise geometric structure, but the quadratic attention cost of T

Feature-Space Smoothing: Certified Robustness of Deep Representations

SafetyDGX agent

arXiv:2601.16200v3 Announce Type: replace-cross Abstract: Modern deep learning models exhibit strong capabilities across diverse applications, yet remain vulnerable to malicious inputs that induce err

Formal Skill: Programmable Runtime Skills for Efficient and Accurate LLM Agents

SafetyDGX agent

arXiv:2605.19604v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly act inside real workspaces, where tools and skills determine whether model reasoning becomes reliable act

Gemini 3.5 announce.

Model ReleasesDGX agent

Google introduced Gemini 3.5, its latest family of models combining frontier intelligence with action capabilities, representing a major leap forward in building more capable, intelligent agents. The

Generalization Bounds of Surrogate Policies for Combinatorial Optimization Problems

Model ReleasesDGX agent

arXiv:2407.17200v3 Announce Type: replace-cross Abstract: Many real-world decision problems require solving, again and again, combinatorial optimization instances drawn from a common distribution. A r

GoTTA be Diverse: Rethinking Memory Policies for Test-Time Adaptation

Model ReleasesDGX agent

arXiv:2605.19890v1 Announce Type: new Abstract: Test-time adaptation (TTA) enables a pre-trained model to adapt online to an unlabeled test stream under distribution shift. While most TTA research foc

HAVEN: Hierarchically Aligned Multimodal Benchmark for Unified Video Understanding

Model ReleasesDGX agent

arXiv:2605.19223v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) exhibit strong performance on standard video tasks, their ability to faithfully summarize and reason over

Hybrid-LoRA: Bridging Full Fine-Tuning and Low-Rank Adaptation for Post-Training

Model ReleasesDGX agent

arXiv:2605.18822v1 Announce Type: cross Abstract: Post-training has become essential for adapting large language models (LLMs) to complex downstream behaviors, including instruction following, prefere

Less Back-and-Forth: A Comparative Study of Structured Prompting

Model ReleasesDGX agent

arXiv:2605.20149v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for open-ended tasks, but underspecified prompts can lead to low-quality answers and additional interacti

Library Hallucinations in LLM-Generated Code: A Risk Analysis Grounded in Developer Queries

Model ReleasesDGX agent

arXiv:2509.22202v3 Announce Type: replace-cross Abstract: Large language models (LLMs) now play a central role in code generation, yet they continue to hallucinate, frequently inventing non-existent l

Lossless Anti-Distillation Sampling

ResearchDGX agent

arXiv:2605.18829v1 Announce Type: new Abstract: Frontier commercial generative models face a growing threat from distillation, whereby a distiller harvests generated responses and trains a competing m

LWiAI Podcast #245 - TML-Interaction, Claude For Legal, Sam Altman on Stand

Model ReleasesDGX agent

This podcast episode covers three main topics: TML-Interaction (likely a new AI model or technical development), the application of Claude AI in legal settings and use cases, and Sam Altman's testimon

MIRO: MultI-Reward cOnditioned pretraining improves T2I quality and efficiency

Model ReleasesDGX agent

arXiv:2510.25897v2 Announce Type: replace Abstract: The default paradigm of post-training text-to-image generators includes post-hoc selection of generated images, and subsequent training with one rew

MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs

Model ReleasesDGX agent

arXiv:2511.14159v2 Announce Type: replace Abstract: Evaluating the robustness of Large Vision-Language Models (LVLMs) is essential for their continued development and responsible deployment in real-wo

Next-Acceleration-Scale Prediction for Autoregressive MRI Reconstruction

Model ReleasesDGX agent

arXiv:2605.19354v1 Announce Type: cross Abstract: MRI reconstruction is an inherently ill-posed inverse problem, since incomplete measurements admit many plausible solutions. This ambiguity becomes mo

Physics-in-the-Loop: A Hybrid Agentic Architecture for Validated CAD Engineering Design

Model ReleasesDGX agent

arXiv:2605.19717v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate Computer-Aided Design (CAD), yet lack physical comprehension required for reliable engineering design. Instead

Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering

SafetyDGX agent

arXiv:2605.19220v1 Announce Type: cross Abstract: Uncertainty Quantification (UQ) is widely regarded as the primary safeguard for deploying Large Language Models (LLMs) in high-stakes domains. However

PromptRad: Knowledge-Enhanced Multi-Label Prompt-Tuning for Low-Resource Radiology Report Labeling

Model ReleasesDGX agent

arXiv:2605.20052v1 Announce Type: cross Abstract: Automatic report labeling facilitates the identification of clinical findings from unstructured text and enables large-scale annotation for medical im

RECIPE: Procedural Planning via Grounding in Instructional Video

Model ReleasesDGX agent

arXiv:2605.19976v1 Announce Type: new Abstract: Visual planning asks a model to generate the remaining steps of a procedure in natural language given a partial video context and a goal. Progress on th

Retrieval-Augmented Generation for Natural Language Processing: A Survey

Model ReleasesDGX agent

arXiv:2407.13193v4 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong empirical performance in various fields, benefiting from their huge amount of parameters that stor

← Previous
1…410411412413414…1059
Next →