AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
23 May 2026

What are the Right Symmetries for Formal Theorem Proving?

SafetyDGX agent

arXiv:2605.22257v1 Announce Type: new Abstract: Formal theorem provers based on large language models (LLMs) are highly sensitive to superficial variations in problem representation: semantically equi

Winner-Take-All bottlenecks enforce disentangled symbolic representations in multi-task learning

ResearchDGX agent

arXiv:2605.22472v1 Announce Type: new Abstract: Winner-take-all (WTA) networks constitute a central circuit motif in cortical networks of the brain. In addition, WTA-like activations are abundant in m

22 May 2026

ACC: Compiling Agent Trajectories for Long-Context Training

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.21850v1 Announce Type: new Abstract: Recent development of agents has renewed demand for long-context reasoning capacity of LLMs. However, training LLMs for this capacity requires costly lo

Action with Visual Primitives

ResearchDGX agent

arXiv:2605.22183v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for generalist robotic manipulation. A common design in current architectures m

An Application-Layer Multi-Modal Covert-Channel Reference Monitor for LLM Agent Egress

AgentsDGX agent

arXiv:2605.20734v1 Announce Type: cross Abstract: A large language model (LLM) agent that sends messages can leak data inside them. Destination allowlists and content scanners do not police whether an

Are -mlx variants equivalent to -nvfp4 for gemma4 and qwen3.6?

Local AiDGX agent

Ollama has added MLX support for Qwen3.5 and Gemma 4 , with both MLX and NVFP4 quantization variants available for these models. There have been issues with broken Qwen3.6 NVFP4 generation on Ollama ,

Assisted Counterspeech Writing at the Crossroads of Hate Speech and Misinformation

ResearchDGX agent

arXiv:2605.22435v1 Announce Type: new Abstract: Hate speech and misinformation frequently co-occur online, amplifying prejudice and polarization. Given their scale, using Large Language Models (LLMs)

Auction-Consensus Algorithm with Learned Bidding Scheme for Multi-Robot Systems

Local AiDGX agent

arXiv:2605.21932v1 Announce Type: new Abstract: Multi-Robot Task Allocation (MRTA) is a central challenge in decentralized multi-agent systems, where teams of robots must cooperatively assign and exec

AutoRPA: Efficient GUI Automation through LLM-Driven Code Synthesis from Interactions

AgentsDGX agent

arXiv:2605.21082v1 Announce Type: new Abstract: Large Language Model (LLM) based agents have demonstrated proficiency in multi-step interactions with graphical user interfaces (GUIs). While most resea

b9276

Local AiDGX agent

llama.cpp build b9276 introduces support for hybrid DNA tokenization with new pre-type dispatching and tokenizer implementations, alongside fixes for VRAM leaks in Multi-Token Prediction (MTP) models

BodyReLux: Temporally Consistent Full-Body Video Relighting

ApplicationsDGX agent

arXiv:2605.21766v1 Announce Type: new Abstract: Being able to relight human performance is a fundamental task for post production and content creation. We present BodyReLux, a subject-specific video d

DeferMem: Query-Time Evidence Distillation via Reinforcement Learning for Long-Term Memory QA

ResearchDGX agent

arXiv:2605.22411v1 Announce Type: new Abstract: Large language model (LLM) agents still struggle with long-term memory question answering, where answer-supporting evidence is often scattered across lo

Demystifying Transition Matching: When and Why It Can Beat Flow Matching

ApplicationsDGX agent

arXiv:2510.17991v3 Announce Type: replace-cross Abstract: Flow Matching (FM) underpins many state-of-the-art generative models, yet recent results indicate that Transition Matching (TM) can achieve hi

Detecting Synthetic Political Narratives in Cross-Platform Social Media Discourse

ResearchDGX agent

arXiv:2605.21540v1 Announce Type: cross Abstract: The proliferation of large language models has introduced a new paradigm of synthetic political communication in which narratives may be generated, se

Diverge to Induce Prompting: Multi-Rationale Induction for Zero-Shot Reasoning

TutorialsDGX agent

arXiv:2602.08028v1 Announce Type: cross Abstract: To address the instability of unguided reasoning paths in standard Chain-of-Thought prompting, recent methods guide large language models (LLMs) by fi

EasyVFX: Frequency-Driven Decoupling for Resource-Efficient VFX Generation

HardwareDGX agent

arXiv:2605.22051v1 Announce Type: new Abstract: Generating high-fidelity visual effects (VFX) typically demands massive datasets and prohibitive computational power due to the intricate coupling of sp

Enabling Regulatory Multi-Agent Collaboration: Architecture, Challenges, and Solutions

AgentsDGX agent

arXiv:2509.09215v2 Announce Type: replace Abstract: Large language models (LLMs)-empowered autonomous agents are transforming both digital and physical environments by enabling adaptive, multi-agent c

EvoIR-Agent: Self-Evolving Image Restoration Agentic System via Experience-Driven Learning

AgentsDGX agent

arXiv:2605.22208v1 Announce Type: new Abstract: Multimodal Large Language Model (MLLM)-driven image restoration agent demonstrates effectiveness in degradation coupling scenarios by flexibly selecting

Exposing Vulnerabilities in Visible-Infrared VLMs: A Unified Geometric Adversarial Framework with Cross-Task Transferability

ApplicationsDGX agent

arXiv:2605.22273v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong performance across diverse multimodal tasks, but their adversarial robustness in visible-infrared (VI

FlyRoute: Self-Evolving Agent Profiling via Data Flywheel for Adaptive Task Routing

SafetyDGX agent

arXiv:2605.22057v1 Announce Type: new Abstract: Enterprise routers assign queries to expert agents, yet deployed profiles stay static while agents evolve (prompts, tools, models), and developers rarel

Great to see @CommonCrawl using and recommending @huggingface Buckets for large constantly evolving training datasets! If you have private m…

IndustryDGX agent

Great to see @CommonCrawl using and recommending @huggingface Buckets for large constantly evolving training datasets! If you have private models or datasets, try it and let us know what you think abo

Improving Viewpoint-Invariance and Temporal Consistency for Action Detection

ResearchDGX agent

arXiv:2605.22695v1 Announce Type: new Abstract: Viewpoint change invariance and action temporal consistency are critical aspects for the effective deployment of human action detection of untrimmed vid

Learning A Unified Risk Map for Autonomous Driving in Partially Observable Environments

AgentsDGX agent

arXiv:2605.22189v1 Announce Type: new Abstract: Occlusion-aware prediction remains a critical challenge in autonomous driving due to the inherent uncertainty of unobserved regions. Existing approaches

LIDSA: Cognitive Arbitration for Signal-Free Autonomous Intersection Management

AgentsDGX agent

arXiv:2605.12321v2 Announce Type: replace Abstract: Large language models (LLMs) show strong potential for Intelligent Transportation Systems (ITS), particularly in tasks requiring situational reasoni

MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering

ResearchDGX agent

arXiv:2605.22269v1 Announce Type: new Abstract: Long streaming video QA remains challenging due to growing visual tokens and limited reasoning length of large language models (LLMs). KV-caching stores

Nicki Minaj template dropped… so I became the baddie. 🔥🚀 @Grok @Imagine

IndustryDGX agent

Nicki Minaj template dropped… so I became the baddie. 🔥🚀 @Grok @Imagine Media Template Drop @SpaceX Starship Baddie (inspired by @NICKIMINAJ) Modeled by @SadieSupaDoge https://grok.com/imagine/templat

Open-source clinical AI is a niche that most days feels like shouting into a PubMed-shaped void. And yet. 5,000 of you are here, building Op…

IndustryDGX agent

Open-source clinical AI is a niche that most days feels like shouting into a PubMed-shaped void. And yet. 5,000 of you are here, building OpenMed in the open with me. Thank you. More models, more data

ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration

HardwareDGX agent

arXiv:2605.22015v1 Announce Type: new Abstract: Diffusion Transformer (DiT) has emerged as a powerful model architecture for generating high-quality images and videos. In the case of video DiT, 3D Spa

PointLLM-R: Enhancing 3D Point Cloud Reasoning via Chain-of-Thought

ApplicationsDGX agent

arXiv:2605.22013v1 Announce Type: new Abstract: Understanding 3D point clouds through language remains a fundamental challenge in computer graphics and visual computing, due to the irregular structure

REACH: Hand Pose Estimation from Room Corners

ResearchDGX agent

arXiv:2605.22231v1 Announce Type: new Abstract: We introduce a novel 3D hand pose estimator that can accurately recover the shape and pose of people's hands in a room from afar, typically from fixed c

Reducing Political Manipulation with Consistency Training

SafetyDGX agent

arXiv:2605.22771v1 Announce Type: new Abstract: Large language models (LLMs) exhibit systematic political bias across a variety of sensitive contexts. We find that LLMs handle counterpart topics from

SENIOR: Efficient Query Selection and Preference-Guided Exploration in Preference-based Reinforcement Learning

SafetyDGX agent

arXiv:2506.14648v2 Announce Type: replace Abstract: Preference-based Reinforcement Learning (PbRL) methods provide a solution to avoid reward engineering by learning reward models based on human prefe

STRUCTSENSE: A Task-Agnostic Agentic Framework for Structured Information Extraction with Human-In-The-Loop Evaluation and Benchmarking

AgentsDGX agent

arXiv:2507.03674v3 Announce Type: replace Abstract: Extracting structured information from scientific literature is critical for accelerating discovery, yet Large Language Models (LLMs) often struggle

Thanks for sharing this. The Huggingface Hub, instead of being considered as a huge dataset, should be considered as a dynamic discovery eng…

IndustryDGX agent

Thanks for sharing this. The Huggingface Hub, instead of being considered as a huge dataset, should be considered as a dynamic discovery engine that can be directly executed and verified. Most models

The Erdős Proof and AI Capabilities

SafetyDGX agent

View the official memo here. An internal model at OpenAI has autonomously disproved a central conjecture in discrete geometry, a mathematical field with applications in cryptography, wireless device c

The state of LLMs in one video #AI

SafetyDGX agent

Gary Marcus discusses the current state of large language models (LLMs), likely covering their capabilities, limitations, and practical applications in AI. The post probably addresses key challenges s

This is the most interesting paper I have read this week. The authors test a wide range of LLMs on a massive dataset of behavioural experime…

SafetyDGX agent

This is the most interesting paper I have read this week. The authors test a wide range of LLMs on a massive dataset of behavioural experiments, with more than 200,000 participants and nearly 26 milli

TriSweep: A Four-Drone Swarm Framework for Electromagnetic Side-Channel Analysis

SafetyDGX agent

arXiv:2605.22709v1 Announce Type: cross Abstract: Electromagnetic (EM) side-channel analysis traditionally assumes a stationary, close-proximity probe - a threat model that underestimates aerial adver

Trump delays AI executive order after tech industry pushback

IndustryDGX agent

U.S. President Donald Trump has delayed the signing of an executive order designed to regulate advanced artificial intelligence models. “I didn’t like certain aspects of it. I postponed it,” Trump tol

Universal CT Representations from Anatomy to Disease Phenotype through Agglomerative Pretraining

SafetyDGX agent

arXiv:2605.21906v1 Announce Type: new Abstract: Computed tomography (CT) is a central to three-dimensional medical imaging, yet CT-based artificial intelligence remains fragmented across task-specific

Value-Gradient Hypothesis of RL for LLMs

SafetyDGX agent

arXiv:2605.21654v1 Announce Type: cross Abstract: Reinforcement learning substantially improves pretrained language models, but it remains understudied why critic-free methods such as PPO and GRPO wor

Video-o3: Native Interleaved Clue Seeking for Long Video Multi-Hop Reasoning

ResearchDGX agent

arXiv:2601.23224v2 Announce Type: replace Abstract: Existing multimodal large language models for long-video understanding predominantly rely on uniform sampling and single-turn inference, limiting th

What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom

AgentsDGX agent

arXiv:2602.01334v2 Announce Type: replace Abstract: Vision tool-use reinforcement learning (RL) can equip vision language models with visual operators such as crop-and-zoom and achieves strong perform

Which Way Did It Move? Diagnosing and Overcoming Directional Motion Blindness in Video-LLMs

Local AiDGX agent

arXiv:2605.22823v1 Announce Type: new Abstract: Video Large Language Models (Video-LLMs) have made rapid progress on temporal video understanding, yet many fail at a basic perceptual primitive: signed

WorldKV: Efficient World Memory with World Retrieval and Compression

HardwareDGX agent

arXiv:2605.22718v1 Announce Type: new Abstract: Autoregressive video diffusion models have enabled real-time, action-conditioned world generation. However, sustaining a persistent world, where revisit

'Would You Want an AI Tutor?' Understanding Stakeholder Perceptions of LLM-based Systems in the Classroom

SafetyDGX agent

arXiv:2503.02885v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have gained traction in educational settings, often framed as virtual tutors or teaching assistants. Following ea

21 May 2026

A New Framework to Analyse the Distributional Robustness of Deep Neural Networks

ApplicationsDGX agent

arXiv:2605.21313v1 Announce Type: new Abstract: Deep neural networks have achieved impressive performance on a variety of tasks, but their brittleness to distributional shifts remains a significant ba

AI-based Prediction of Independent Construction Safety Outcomes from Universal Attributes

SafetyDGX agent

arXiv:1908.05972v3 Announce Type: replace Abstract: This paper significantly improves on, and finishes to validate, an approach proposed in previous research in which safety outcomes were predicted fr

AIGaitor: Privacy-preserving and cloud-free motion analysis for everyone, using edge computing

Local AiDGX agent

arXiv:2605.21421v1 Announce Type: new Abstract: Motion capture is the gold standard for measuring human movement, but clinical use remains limited by cost, technical complexity, and privacy concerns.

Amazon Nova Act is now HIPAA eligible

AgentsDGX agent

Amazon Nova Act is a service for building and managing AI agents that automate browser-based UI workflows, powered by a custom Nova 2 Lite model. It is particularly valuable for industries with fragme

Automated Byzantine-Resilient Clustered Decentralized Federated Learning for Battery Intelligence in Connected EVs

SafetyDGX agent

arXiv:2605.21115v1 Announce Type: cross Abstract: Federated learning (FL) has emerged as a promising paradigm for managing electric vehicle (EV) battery data in intelligent transportation systems (ITS

Beyond Words: Multimodal LLM Knows When to Speak

AgentsDGX agent

arXiv:2505.14654v2 Announce Type: replace-cross Abstract: Chatbots via large language models (LLMs) generate fluent responses but often struggle with when to speak, especially for brief, timely listen

C^2FG: Control Classifier-Free Guidance via Score Discrepancy Analysis

ResearchDGX agent

arXiv:2603.08155v3 Announce Type: replace Abstract: Classifier-Free Guidance (CFG) is a cornerstone of modern conditional diffusion models, yet its reliance on the fixed or heuristic dynamic guidance

Can Microcanonical Langevin Dynamics Leverage Mini-Batch Gradient Noise?

SafetyDGX agent

arXiv:2602.06500v2 Announce Type: replace Abstract: Scaling inference methods such as Markov chain Monte Carlo to high-dimensional models remains a central challenge in Bayesian deep learning. A promi

CHEM: Estimating and Understanding Hallucinations in Deep Learning for Image Processing

SafetyDGX agent

arXiv:2512.09806v2 Announce Type: replace Abstract: Deep learning-based methods have recently achieved significant success in image reconstruction problems. However, challenges have emerged, as these

Creé mi propio agente de IA local centrado en Ollama y la ejecución de tareas prácticas.

Local AiDGX agent

This post describes a user's experience building a local AI agent using Ollama, an open-source tool for running large language models locally. The article likely covers the practical implementation st

Dynamic TMoE: A Drift-Aware Dynamic Mixture of Experts Framework for Non-Stationary Time Series Forecasting

ResearchDGX agent

arXiv:2605.20678v1 Announce Type: new Abstract: Non-stationary time series forecasting is challenged by evolving distribution shifts that static models struggle to capture. While Mixture-of-Experts (M

EDA Market Primer - Market Dynamics, Cadence, Synopsys, Siemens, China EDA Rise

HardwareDGX agent

EDA Market size, Share, Business Models, Drivers, Changing Customer Base, Competitive Dynamics Across Synopsys, Cadence, and Siemens, China EDA, IP, Hardware, CoT, Lock-In Economics, Disruptive Forces

Efficient Banzhaf-Based Data Valuation for k-Nearest Neighbors Classification

ApplicationsDGX agent

arXiv:2605.21033v1 Announce Type: new Abstract: Data valuation, the task of quantifying the contribution of individual data points to model performance, has emerged as a fundamental challenge in machi

END: Early Noise Dropping for Efficient and Effective Context Denoising

ResearchDGX agent

arXiv:2502.18915v3 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable performance across a wide range of natural language processing tasks. However, they are of

← Previous
1…767768769770771…1018
Next →