AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,595 results
21 May 2026

Decomposing Subject-Driven Image Generation via Intermediate Structural Prediction

Model ReleasesDGX agent

arXiv:2605.20807v1 Announce Type: new Abstract: Subject-driven text-to-image generation still struggles to preserve high-frequency identity details such as logos, patterns, and text. Existing methods

Deep Neural Networks as Discrete Dynamical Systems: Implications for Physics-Informed Learning

Model ReleasesDGX agent

arXiv:2601.00473v3 Announce Type: replace Abstract: We revisit the analogy between feed-forward deep neural networks (DNNs) and discrete dynamical systems derived from neural integral equations and th

DEL: Digit Entropy Loss for Numerical Learning of Large Language Models

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.20369v1 Announce Type: new Abstract: Number prediction stands as a fundamental capability of large language models (LLMs) in mathematical problem-solving and code generation. The widely ado

Deltaynamics: Language-Based Representation for Inferring Rigid-Body Dynamics From Videos

Model ReleasesDGX agent

arXiv:2605.20576v1 Announce Type: new Abstract: Inferring rigid-body physical states and properties from monocular videos is a fundamental step toward physics-based perception and simulation. Existing

Diagnosing Overhead in Dispatch Operations: Cross-architecture Observatory

Model ReleasesDGX agent

arXiv:2605.20982v1 Announce Type: cross Abstract: AlltoAll dispatch is the dominant bottleneck of MoE expert parallelism, and the interconnect community has responded with four families of mitigations

DiMextsuperscript{3}: Bridging Multilingual and Multimodal Models via Direction- and Magnitude-Aware Merging

Model ReleasesDGX agent

arXiv:2605.12960v2 Announce Type: replace Abstract: Towards more general and human-like intelligence, large language models should seamlessly integrate both multilingual and multimodal capabilities; h

DISC: Decoupling Instruction from State-Conditioned Control via Policy Generation

Model ReleasesDGX agent

arXiv:2605.20856v1 Announce Type: cross Abstract: Language-conditioned manipulation policies typically process instructions and observations through shared network parameters. This task-state entangle

DIVE: Embedding Compression via Self-Limiting Gradient Updates

Model ReleasesDGX agent

arXiv:2605.20689v1 Announce Type: new Abstract: High-dimensional embeddings from large language models impose significant storage and computational costs on vector search systems. Recent embedding com

Divide et Calibra: Multiclass Local Calibration via Vector Quantization

Model ReleasesDGX agent

arXiv:2605.21060v1 Announce Type: new Abstract: Accurate and well-calibrated Machine Learning (ML) models are mandatory in high-stakes settings, yet effective multiclass calibration remains challengin

Do LLMs Know What Luxembourgish Borrows? Probing Lexical Neology in Low-Resource Multilingual Models

Model ReleasesDGX agent

arXiv:2605.21227v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for writing assistance in small contact languages, yet it is unclear whether they respect community n

Do Vision--Language Models Understand 3D Scenes or Just Catalogue Objects?

Model ReleasesDGX agent

arXiv:2605.20448v1 Announce Type: new Abstract: Vision--language models reliably name objects in a scene, but do they represent the 3D layout those objects inhabit? We introduce a 3,034-sample human-c

DriveMA: Rethinking Language Interfaces in Driving VLAs with One-Step Meta-Actions

Model ReleasesDGX agent

arXiv:2605.21273v1 Announce Type: new Abstract: Driving Vision-Language-Action Models (Driving VLAs) commonly introduce natural-language reasoning as an intermediate interface for end-to-end planning,

DrugRAG: Enhancing Pharmacy LLM Performance Through A Novel Retrieval-Augmented Generation Pipeline

Model ReleasesDGX agent

arXiv:2512.14896v2 Announce Type: replace Abstract: In our study, we evaluated large language model (LLM) performance on pharmacy licensure-style question-answering tasks and developed an external kno

Dynamic Video Generation: Shaping Video Generation Across Time and Space

Model ReleasesDGX agent

arXiv:2605.21042v1 Announce Type: new Abstract: Diffusion models have achieved impressive performance in video generation, but their iterative denoising process remains computationally expensive due t

DySink: Dynamic Frame Sinks for Autoregressive Long Video Generation

Model ReleasesDGX agent

arXiv:2605.21028v1 Announce Type: new Abstract: Autoregressive long video generation often adopts bounded-memory streaming for efficiency, typically combining local windows for short-term continuity w

ECUAS{n}: A family of metrics for principled evaluation of uncertainty-augmented systems

Model ReleasesDGX agent

arXiv:2605.20490v1 Announce Type: cross Abstract: In high-stakes automated decision-making, access to predictive uncertainty is essential for enabling users -- human or downstream systems -- to accept

Enhanced Reinforcement Learning-based Process Synthesis via Quantum Computing

Model ReleasesDGX agent

arXiv:2605.21213v1 Announce Type: cross Abstract: In this work, we present quantum reinforcement learning (RL) as a solution strategy for process synthesis problems. Building on our prior work, we dev

Evolutionary Generation of Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2602.06511v3 Announce Type: replace Abstract: Large language model (LLM)-based multi-agent systems (MAS) show strong promise for complex reasoning, planning, and tool-augmented tasks, but design

Explainability Methods for Hardware Trojan Detection: A Systematic Comparison

Model ReleasesDGX agent

arXiv:2601.18696v4 Announce Type: replace Abstract: Hardware trojans are malicious circuits which compromise the functionality and security of an integrated circuit (IC). These circuits are manufactur

Exploring Deep Learning and Ultra-Widefield Imaging for Diabetic Retinopathy and Macular Edema

Model ReleasesDGX agent

arXiv:2603.08235v2 Announce Type: replace Abstract: Diabetic retinopathy (DR) and diabetic macular edema (DME) are leading causes of preventable blindness among working-age adults. Traditional approac

FAIR-Pruner: A Flexible Framework for Automatic Layer-Wise Pruning via Tolerance of Difference

Model ReleasesDGX agent

arXiv:2508.02291v3 Announce Type: replace Abstract: Structured pruning is a standard tool for compressing deep neural networks, but its practical performance depends on how sparsity is allocated acros

FBOS-RL: Feedback-Driven Bi-Objective Synergistic Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.20256v1 Announce Type: new Abstract: Reinforcement learning has become a cornerstone for aligning and unlocking the reasoning capabilities of large-scale models. At its core, the training l

FedCoE: Bridging Generalization and Personalization via Federated Coordinated Dual-level MoEs

Model ReleasesDGX agent

arXiv:2605.21264v1 Announce Type: new Abstract: Federated Learning (FL) has emerged as a promising paradigm for privacy-preserving distributed learning. However, existing FL methods face a fundamental

FedCritic: Serverless Federated Critic Learning-based Resource Allocation for Multi-Cell OFDMA in 6G

Model ReleasesDGX agent

arXiv:2605.21418v1 Announce Type: cross Abstract: In sixth-generation (6G) ultra-dense networks, aggressive frequency reuse amplifies inter-cell interference (ICI), making multi-cell orthogonal freque

Federal records: Grok was utilized in only 3 of 400+ publicly identified federal AI use cases in 2025, behind 234 for ChatGPT, 33 for Gemini, 26 for Claude (Reuters)

Model ReleasesDGX agent

Reuters: Federal records: Grok was utilized in only 3 of 400+ publicly identified federal AI use cases in 2025, behind 234 for ChatGPT, 33 for Gemini, 26 for Claude — SpaceX's initial public offering

Federated LoRA Fine-Tuning for LLMs via Collaborative Alignment

Model ReleasesDGX agent

arXiv:2605.21217v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has emerged as a powerful tool for parameter-efficient fine-tuning of large language models (LLMs). This paper studies LoRA

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy

Model ReleasesDGX agent

arXiv:2605.20965v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have shown remarkable performance on a wide range of vision-language tasks. Despite this progress, they are still p

Findings of the Counter Turing Test: AI-Generated Text Detection

Model ReleasesDGX agent

arXiv:2605.20761v1 Announce Type: new Abstract: The rapid proliferation of AI-generated text has introduced significant challenges in maintaining the integrity of digital content. Advanced generative

Fine-grained Claim-level RAG Benchmark for Law

Model ReleasesDGX agent

arXiv:2605.21071v1 Announce Type: new Abstract: The rapid progress of large language models (LLMs) is shifting semantic search toward a question-answering paradigm, where users ask questions and LLMs

Free-Grained Hierarchical Visual Recognition

Model ReleasesDGX agent

arXiv:2510.14737v3 Announce Type: replace Abstract: Hierarchical image recognition seeks to predict class labels along a semantic taxonomy, from broad categories to specific ones, typically under the

FT-Dojo: Towards Autonomous LLM Fine-Tuning with Language Agents

Model ReleasesDGX agent

arXiv:2603.01712v2 Announce Type: replace-cross Abstract: Fine-tuning large language models for vertical domains remains labor-intensive, requiring practitioners to curate data, configure training, an

FullFlow: Upgrading Text-to-Image Flow Matching Models for Bidirectional Vision--Language Generation

Model ReleasesDGX agent

arXiv:2605.20316v1 Announce Type: new Abstract: Modern text-to-image diffusion models encode rich visual priors, but expose them only through one-way text-conditioned generation. Existing unified visi

Gated Normalization Removal and Scale Anchoring in Pre-Norm Transformers

Model ReleasesDGX agent

arXiv:2602.10408v2 Announce Type: replace-cross Abstract: Normalization layers are standard in transformers, but it is not clear whether their sample-dependent computations are necessary throughout bo

GenAI-Driven Threat Detection with Microsoft Security Copilot

Model ReleasesDGX agent

arXiv:2605.20896v1 Announce Type: cross Abstract: Defending against today's increasingly sophisticated cyberattacks requires security analysts to continuously translate evolving attacker tradecraft in

Geometry-Lite: Interpretable Safety Probing via Layer-Wise Margin Geometry

Model ReleasesDGX agent

arXiv:2605.20241v1 Announce Type: cross Abstract: Prompt-level safety probes for large language models use hidden-state representations to separate safe from unsafe prompts, but strong average detecti

Google just revealed Omni, personalized cross-device intelligence, and Spark agents at I/O 2025. I sat down with CEO Sundar Pichai to figure…

Model ReleasesDGX agent

Google just revealed Omni, personalized cross-device intelligence, and Spark agents at I/O 2025. I sat down with CEO Sundar Pichai to figure out what comes next: 1:46 Omni: 'Nano Banana for video' 4:5

GradPower: Powering Gradients for Faster Language Model Pre-Training

Model ReleasesDGX agent

arXiv:2505.24275v3 Announce Type: replace Abstract: We propose GradPower, a lightweight gradient-transformation technique for accelerating language model pre-training. Given a gradient vector g=(g_i)_

GraphRAG on Consumer Hardware: Benchmarking Local LLMs for Healthcare EHR Schema Retrieval

Model ReleasesDGX agent

arXiv:2605.20815v1 Announce Type: new Abstract: Graph-based Retrieval Augmented Generation (GraphRAG) extends retrieval-augmented generation to support structured reasoning over complex corpora, but i

Group-Algebraic Tensors: Provably-optimal Equivariant Learning and Physical Symmetry Discovery

Model ReleasesDGX agent

arXiv:2605.20440v1 Announce Type: new Abstract: We introduce the star_G tensor algebra, in which any finite group G defines the multiplication rule, making equivariance an intrinsic algebraic property

Hack-Verifiable Environments: Towards Evaluating Reward Hacking at Scale

Model ReleasesDGX agent

arXiv:2605.20744v1 Announce Type: new Abstract: Aligning autonomous agents with human intent remains a central challenge in modern AI. A key manifestation of this challenge is reward hacking, whereby

HalluCXR: Benchmarking and Mitigating Hallucinations in Medical Vision-Language Models for Chest Radiograph Interpretation

Model ReleasesDGX agent

arXiv:2605.20469v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used for medical image interpretation, yet they frequently hallucinate, generating clinically plausible b

How Much Online RL is Enough? Informative Rollouts for Offline Preference Optimization in RLVR

Model ReleasesDGX agent

arXiv:2605.21266v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as a powerful paradigm for reasoning in language models, with GRPO as its primary exam

HRM-Text: Efficient Pretraining Beyond Scaling

Model ReleasesDGX agent

arXiv:2605.20613v1 Announce Type: new Abstract: The current pretraining paradigm for large language models relies on massive compute and internet-scale raw text, creating a significant barrier to foun

Hugging Face just released LeRobot Humanoid An open-source, low-cost (~$2.5k), 3D-printed humanoid built for robot learning and not just dem…

Model ReleasesDGX agent

Hugging Face just released LeRobot Humanoid An open-source, low-cost (~$2.5k), 3D-printed humanoid built for robot learning and not just demos. What’s cool is it’s a full stack release: • hardware + C

@huggingface's @LeRobotHF team announced LeRobot Humanoid, an open-source bipedal robot platform built for roughly $2,500 using mostly 3D-pr…

Model ReleasesDGX agent

@huggingface's @LeRobotHF team announced LeRobot Humanoid, an open-source bipedal robot platform built for roughly $2,500 using mostly 3D-printed and off-the-shelf parts: The release provides complete

Humanoid Whole-Body Manipulation via Active Spatial Brain and Generalizable Action Cerebellum

Model ReleasesDGX agent

arXiv:2605.21133v1 Announce Type: new Abstract: In this paper, we explore spatial-aware humanoid whole-body manipulation task. Compared with tabletop settings, this task poses two key challenges: 1) S

Hyper-V2X: Hypernetworks for Estimating Epistemic and Aleatoric Uncertainty in Cooperative Bird's-Eye-View Semantic Segmentation

Model ReleasesDGX agent

arXiv:2605.21309v1 Announce Type: new Abstract: Cooperative perception enabled by Vehicle-to-Everything (V2X) communication enhances autonomous driving safety by creating a unified environmental repre

I released the first alpha of Datasette Agent - a conversational AI assistant for Datasette that can answer questions about data in SQLite d…

Model ReleasesDGX agent

I released the first alpha of Datasette Agent - a conversational AI assistant for Datasette that can answer questions about data in SQLite databases, and can be extended with plugins to add extra tool

If this is true, using the best public estimates we have of LLM resource use, solving this Erdos problem took 0.6–6.3 kWh of electricity and…

Model ReleasesDGX agent

If this is true, using the best public estimates we have of LLM resource use, solving this Erdos problem took 0.6–6.3 kWh of electricity and about 3–31 liters of water. So that is less than three almo

Improving Quantized Model Performance in Qualitative Analysis with Multi-Pass Prompt Verification

Model ReleasesDGX agent

arXiv:2605.20193v1 Announce Type: new Abstract: Quantized Large Language Models (LLMs) are used more often in qualitative analysis because they run fast and need fewer computing resources. This study

In the next version of Claude Code: run /usage to see a breakdown of which Skills, Agents, MCPs, and Plugins are using your tokens CLI today…

Model ReleasesDGX agent

The next version of Claude Code will introduce a `/usage` command that provides a detailed breakdown of token consumption across different components including Skills, Agents, MCPs (Model Context Prot

InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling

Model ReleasesDGX agent

arXiv:2508.08636v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized artificial intelligence by enabling complex reasoning capabilities. While recent advancements in re

JFAA: Technical Report for the EPIC-KITCHENS-100 Action Anticipation Challenge at EgoVis 2026

Model ReleasesDGX agent

arXiv:2605.20904v1 Announce Type: new Abstract: We propose JFAA, a JEPA-based Future Action Anticipation method for the EPIC-KITCHENS-100 (EK-100) Action Anticipation task. Inspired by the representat

JobArabi: An Arabic Corpus and Analysis of Job Announcements from Social Media

Model ReleasesDGX agent

arXiv:2605.20960v1 Announce Type: new Abstract: This paper introduces JobArabi, a large-scale corpus of Arabic job announcements collected from social media between January 2024 and October 2025. The

JUDO: A Juxtaposed Domain-Oriented Multimodal Reasoner for Industrial Anomaly QA

Model ReleasesDGX agent

arXiv:2605.20284v1 Announce Type: new Abstract: Industrial anomaly detection has been significantly advanced by Large Multimodal Models (LMMs), enabling diverse human instructions beyond detection, pa

LAION-C: An Out-of-Distribution Benchmark for Web-Scale Vision Models

Model ReleasesDGX agent

arXiv:2506.16950v2 Announce Type: replace Abstract: Out-of-distribution (OOD) robustness is a desired property of computer vision models. Improving model robustness requires high-quality signals from

LamPO: A Lambda Style Policy Optimization for Reasoning Language Models

Model ReleasesDGX agent

arXiv:2605.21235v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving reasoning language models on tasks such as mathemat

Large-Step Training Dynamics of a Two-Factor Linear Transformer Model

Model ReleasesDGX agent

arXiv:2605.21292v1 Announce Type: cross Abstract: Gradient-flow analyses show that simplified linear transformers can learn the in-context linear-regression algorithm, but they do not explain the fini

Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search

Model ReleasesDGX agent

arXiv:2605.20244v1 Announce Type: cross Abstract: We present Lean Refactor, a plug-and-play retrieval-augmented agentic framework for multi-objective, controllable, and version-robust refactoring of L

LEAP: A closed-loop framework for perovskite precursor additive discovery

Model ReleasesDGX agent

arXiv:2605.20242v1 Announce Type: new Abstract: Efficient discovery of precursor additives is essential for improving the performance of perovskite solar cells, yet the large chemical space makes conv

← Previous
1…223224225226227…377
Next →