AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
2 Jun 2026

TERRA: Task-Embedded Reasoning and Representation Architecture for Cross-Domain Applications

ResearchDGX agent

arXiv:2606.01520v1 Announce Type: new Abstract: A single action-conditioned latent predictive architecture can in principle be trained on the structured state of a driving scene, a robot workspace, or

The art and science of hyperparameter optimization on Amazon Nova Forge

TutorialsDGX agent

Fine-tuning for domain-specific tasks means improving performance in one area without degrading the model’s general capabilities, and getting that balance right is harder than it looks. This post walk

The Social Cost of Intelligence: Emergence, Propagation, and Amplification of Stereotypical Bias in Multi-Agent Systems

SafetyDGX agent

arXiv:2510.10943v2 Announce Type: replace-cross Abstract: Bias in large language models (LLMs) remains a persistent challenge, often leading to stereotyping and unfair treatment across social groups.

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TIDES: Time-Derivative Event Simulation via Deformable Reconstruction

TutorialsDGX agent

arXiv:2606.02058v1 Announce Type: new Abstract: Event cameras emit asynchronous events in response to environmental appearance changes. The scarcity of real-world event datasets makes simulation essen

Toward Responsible and Epistemically Grounded Multilingual LLMs for Computational Social Science and Humanities

SafetyDGX agent

arXiv:2606.00596v1 Announce Type: new Abstract: Large language models have rapidly evolved in multilingual competence and reasoning capacity, enabling their integration into Social Sciences and Humani

Toward Robust In-Context Learning: Leveraging Out-of-distribution Proxies for Target Inaccessible Demonstration Retrieval

TutorialsDGX agent

arXiv:2606.00014v1 Announce Type: cross Abstract: Although studies have demonstrated that Large Language Models (LLMs) can perform well on Out-of-Distribution (OOD) tasks, their advantage tends to dim

Towards Precise Intent-Aligned VLA Aerial Navigation via Expert-Guided GRPO

SafetyDGX agent

arXiv:2606.02313v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising end-to-end paradigm for unmanned aerial vehicles (UAVs) to accomplish complex tasks specified by f

TrafficClaw: A Generalizable LLM Agent in the Unified Physical Environment for Urban Traffic Control

Local AiDGX agent

arXiv:2604.17456v2 Announce Type: replace Abstract: Large language model (LLM) agents have shown strong capabilities in long-horizon reasoning, tool use, and decision-making in digital environments, y

TrafficRAG: A Multimodal RAG Framework for Traffic Accident Liability Determination

ApplicationsDGX agent

arXiv:2606.01737v1 Announce Type: new Abstract: Traffic accident liability analysis is a critical yet challenging task in intelligent transportation and legal assistance. Existing methods often suffer

Training Prompt Matters: State-Adaptive Optimization for Robust Fine-Tuning

ResearchDGX agent

arXiv:2606.01967v1 Announce Type: new Abstract: While prompt engineering is instrumental in maximizing the capabilities of Large Language Models (LLMs) during inference, the role of prompts during tra

TRON: Targeted Rule-Verifiable Online Environments for Visual Reasoning RL

ResearchDGX agent

arXiv:2606.01599v1 Announce Type: new Abstract: Reinforcement learning (RL) for visual reasoning needs scalable, verifiable, and controllable training signals. Existing visual RL post-training trains

Ultra Diffusion Poser: Diffusion-Based Human Motion Tracking From Sparse Inertial Sensors and Ranging-Based Between-Sensor Distances

SafetyDGX agent

arXiv:2606.02153v1 Announce Type: new Abstract: Methods using inertial measurement units (IMUs) provide a wearable alternative to camera-based motion capture. To mitigate drift from inertial signals,

UME: A Unified Meta-Generalization Framework for Cross-Domain ETA

ApplicationsDGX agent

arXiv:2606.00979v1 Announce Type: new Abstract: Accurate Estimated Time of Arrival (ETA) prediction on checkout page is crucial in instant logistics for enhancing user satisfaction, optimizing dispatc

Vegas: Self-Speculative Decoding with Verification-Guided Sparse Attention

ResearchDGX agent

arXiv:2602.07223v2 Announce Type: replace Abstract: Long-context large language model (LLM) inference has become the norm for today's AI applications. However, it is severely bottlenecked by the incre

Versatile Framework with Semantic and Structural guidance for Image Reconstruction from Brain Activity

ResearchDGX agent

arXiv:2606.00121v1 Announce Type: cross Abstract: Reconstructing visual stimuli from brain recordings has been a meaningful and challenging task in brain decoding. Especially, the achievement of preci

VISReg: Variance-Invariance-Sketching Regularization for JEPA training

ResearchDGX agent

arXiv:2606.02572v1 Announce Type: new Abstract: Self-supervised learning methods prevent embedding collapse via modeling heuristics or explicit regularization of the embedding space. Among the latter,

WaveFilter: Enhancing the Long-Context Capability of Diffusion LLMs via Wavelet-Guided KV Cache Filtering

ResearchDGX agent

arXiv:2606.00724v1 Announce Type: cross Abstract: Diffusion Large Language Models (DLMs) have demonstrated significant advantages across various tasks. However, constrained by their multi-step iterati

What’s new in Microsoft Foundry | Build Edition

IndustryDGX agent

Microsoft Build 2026 brings a major set of Microsoft Foundry updates for developers building agents: hosted runtimes, Toolboxes, memory, Voice Live, Foundry IQ, new models, managed compute, and trust,

When Rating Scales Fall Short: LLM-Assisted Discovery of ADHD Signals in Turkish Teacher Narratives

ResearchDGX agent

arXiv:2606.02509v1 Announce Type: new Abstract: Attention Deficit Hyperactivity Disorder (ADHD) is one of the most common neurodevelopmental disorders in childhood, and its diagnosis relies on assessm

Who Evaluates AI's Social Impacts? Mapping Coverage and Gaps in First and Third Party Evaluations

SafetyDGX agent

arXiv:2511.05613v2 Announce Type: replace-cross Abstract: Foundation models are increasingly central to high-stakes AI systems, and governance frameworks now depend on evaluations to assess their risk

1 Jun 2026

An Organization-Scoped LLM Agent Runtime Architecture for Regulated Cybersecurity Operations

Local AiDGX agent

arXiv:2605.30604v1 Announce Type: cross Abstract: Regulated cybersecurity workflows lack a runtime substrate that enforces organization-level scope across retrieval, tool calls, memory, findings, repo

b9451

Local AiDGX agent

b9451 is the latest release of llama.cpp, published on June 1, 2026 . Llama.cpp is an LLM inference project implemented in C/C++ that enables efficient local execution of large language models. This r

b9453

Local AiDGX agent

llama.cpp release b9453 added support for EXAONE 4.5 model implementations with vision capabilities, including markers, projector paths, and routing through a Qwen2.5-VL-style encode path with window

BackSplit: The Importance of Sub-dividing the Background in Biomedical Lesion Segmentation

ResearchDGX agent

arXiv:2511.19394v2 Announce Type: replace Abstract: Segmenting small lesions in medical images remains notoriously difficult. Most prior work tackles this challenge by either designing better architec

BAT: Better Audio Transformer Guided by Convex Gated Probing

TutorialsDGX agent

arXiv:2602.16305v2 Announce Type: replace-cross Abstract: Probing is widely adopted in computer vision to faithfully evaluate self-supervised learning (SSL) embeddings, as finetuning may misrepresent

Beyond LLMs: Why Scalable Enterprise AI Adoption Depends on Agent Logic

AgentsDGX agent

Enterprise AI adoption at scale requires moving beyond large language models to implement agent logic systems that can handle complex reasoning, planning, and decision-making autonomously. Agent-based

CacheProbe: Auditing Prompt Cache Isolation in Gateway APIs

ResearchDGX agent

arXiv:2605.30613v1 Announce Type: cross Abstract: Over the past year, prompt caching in Large Language Models (LLMs) has become increasingly more popular across inference APIs. Prompt caching helps sa

Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation

ResearchDGX agent

arXiv:2510.22067v3 Announce Type: replace Abstract: Vision language models (VLMs) often generate hallucination, i.e., content that cannot be substantiated by either textual or visual inputs. Prior wor

Chain-of-Thought and Compressed Looped Transformers: A Memory-Budget Separation

ResearchDGX agent

arXiv:2605.30757v1 Announce Type: new Abstract: Chain-of-thought prompting and looped Transformers both give a fixed model more test-time computation, but they differ in what they remember. Chain-of-t

Chatterbox-Flash: Prior-Calibrated Block Diffusion for Streaming Zero-Shot TTS

ResearchDGX agent

arXiv:2605.30748v1 Announce Type: cross Abstract: We present Chatterbox-Flash, a zero-shot text-to-speech model obtained by fine-tuning a pretrained autoregressive TTS decoder into a block-diffusion d

ConSensus: Multi-Agent Collaboration for Multimodal Sensing

SafetyDGX agent

arXiv:2601.06453v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly grounded in sensor data to perceive and reason about human physiology and the physical world. However,

Cross-Lingual Steering for Figurative Language Generation

ResearchDGX agent

arXiv:2605.30443v1 Announce Type: new Abstract: Multilingual large language models can generate figurative language, but whether the internal signals driving this behavior are language-specific or reu

Cross-Modal Attention Calibration for LVLM Hallucination Mitigation

SafetyDGX agent

arXiv:2501.01926v3 Announce Type: replace-cross Abstract: Large vision-language models (LVLMs) have shown remarkable capabilities in visual-language understanding. Despite their success, LVLMs still s

Cross-Modal Clinical Knowledge Integration for Mammography Report Generation

ApplicationsDGX agent

arXiv:2605.31093v1 Announce Type: new Abstract: Breast cancer is a major global health concern, and mammography screening plays a central role in early detection. The large volume of screening examina

D^3: Dynamic Directional Graph-Constrained Data Scheduling for LLM Training

ApplicationsDGX agent

arXiv:2605.31164v1 Announce Type: cross Abstract: Training data plays a central role in large language models (LLMs) optimization, motivating extensive research on data scheduling strategies. Most exi

Decoding the Surgical Scene: A Scoping Review of Scene Graphs in Surgery

SafetyDGX agent

arXiv:2509.20941v2 Announce Type: replace Abstract: As surgical AI transitions from pixel-level detection to complex reasoning, Scene Graphs (SGs) offer the structured, relational representations nece

Deterministic Inference across Tensor Parallel Sizes That Eliminates Training-Inference Mismatch

HardwareDGX agent

arXiv:2511.17826v2 Announce Type: replace-cross Abstract: Deterministic inference is increasingly critical for large language model (LLM) applications such as LLM-as-a-judge evaluation, multi-agent sy

Don't Fool Me Twice: Adapting to Adversity in the Wild with Experience-Driven Reasoning

AgentsDGX agent

arXiv:2605.31119v1 Announce Type: cross Abstract: In robotics, dangers and adversity modes are often embodiment-specific and relative to each agent. A frontier of autonomous mobile robotics is to enab

DRIFT: Decoupled Rollouts and Importance-Weighted Fine-Tuning for Efficient Multi-Turn Optimization

SafetyDGX agent

arXiv:2605.31455v1 Announce Type: cross Abstract: Large language models are increasingly deployed in multi-turn interactive settings where users or environments can iteratively provide lightweight fee

Early Prediction of Future Behavioral Strategy from Process Traces

ResearchDGX agent

arXiv:2605.30550v1 Announce Type: new Abstract: Adaptive systems often need to make task-specific decisions about people from limited evidence: a tutor may need to anticipate how a learner will approa

Efficient and Uncertainty-Aware Diffusion Framework for Offline-to-Online Reinforcement Learning

SafetyDGX agent

arXiv:2605.30776v1 Announce Type: new Abstract: Offline-to-Online Reinforcement Learning (O2O-RL) leverages an offline, pre-trained policy to minimize costly online interactions. Although data-efficie

Efficient Diffusion LLMs via Temporal-Spatial Parallel Decoding and Confidence Extrapolation

ResearchDGX agent

arXiv:2605.30753v1 Announce Type: new Abstract: Diffusion-based large language models (dLLMs) support parallel text generation via iterative denoising, yet inference remains latency-heavy because many

EMCEE: Improving Multilingual Capability of LLMs via Bridging Knowledge and Reasoning with Extracted Synthetic Multilingual Context

ResearchDGX agent

arXiv:2503.05846v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have achieved impressive progress across a wide range of tasks, yet their heavy reliance on English-centric train

Equivariant Latent Alignment via Flow Matching under Group Symmetries

SafetyDGX agent

arXiv:2605.30705v1 Announce Type: new Abstract: Geometry-aware generative models and novel view synthesis approaches have shown strong potential in visual fidelity and consistency. In parallel, equiva

EvoGens: A Population-Based Heuristic Search Framework for Scientific Idea Generation

ResearchDGX agent

arXiv:2605.30961v1 Announce Type: new Abstract: Generating novel research ideas is fundamental to scientific progress. While Large Language Models (LLMs) show promise in assisting this process, existi

Exploiting Chordal Sparsity for Globally Optimal Estimation with Factor Graphs

SafetyDGX agent

arXiv:2605.30617v1 Announce Type: new Abstract: Robust and efficient state estimation is crucial for perception, navigation, and control in robotics. State estimation problems are conveniently modeled

Extending MCP support for Amazon Bedrock AgentCore Gateway

AgentsDGX agent

While deploying Model Context Protocol (MCP) servers in production, enterprises need fine-grained access control across servers, observability into which teams use which tools, security guarantees aga

Forgetting Has Neighbors: Localized Collateral Forgetting in Machine Unlearning

Local AiDGX agent

arXiv:2605.31317v1 Announce Type: new Abstract: Machine unlearning aims to remove the influence of selected training examples without full retraining. Standard evaluations often summarize unlearning q

Functional MRI Time Series Generation via Wavelet-Based Image Transform and Spectral Flow Matching for Brain Disorder Identification

Local AiDGX agent

arXiv:2605.30387v1 Announce Type: cross Abstract: Functional Magnetic Resonance Imaging (fMRI) provides non-invasive access to dynamic brain activity by measuring blood oxygen level-dependent (BOLD) s

Generative Drifting is Secretly Score Matching: a Spectral and Variational Perspective

TutorialsDGX agent

arXiv:2603.09936v2 Announce Type: replace Abstract: Generative Modeling via Drifting~itep{deng2026drifting} has recently achieved state-of-the-art one-step image generation through a kernel-based drif

Getting Started: 1. Update ComfyUI to the latest version [] 2. Open the Template library, then search for TripoSplat 3. Follow the note in t…

Local AiDGX agent

Getting Started: 1. Update ComfyUI to the latest version [] 2. Open the Template library, then search for TripoSplat 3. Follow the note in the workflow to download models 4. Load an image, then run th

GRKV: Global Regression for Training-Free KV Cache Compression in Long-Context LLMs

ResearchDGX agent

arXiv:2605.31105v1 Announce Type: new Abstract: Large language models (LLMs) with extended context lengths rely on the key-value (KV) cache to support attention over prior tokens. However, maintaining

HADT: A Heterogeneous Multi-Agent Differential Transformer for Autonomous Earth Observation Satellite Cluster

AgentsDGX agent

arXiv:2605.31023v1 Announce Type: new Abstract: This work addresses the problem of autonomous resource management in heterogeneous satellite cluster conducting Earth Observation (EO) missions includin

Haptic Sorter: A Unified Planning Framework for Online Shape Estimation and Real-Time Pose Inference

TutorialsDGX agent

arXiv:2605.31352v1 Announce Type: new Abstract: Robotics manipulation usually assumes that the shape and pose of the object are known to the robot prior to motion planning. However, precise geometric

Hermes Agent is now natively supported on @Windows

AgentsDGX agent

Nous Research announced native support for Hermes Agent on Windows, expanding the availability of their Hermes model to Windows-based systems. This development enables Windows users to run Hermes Agen

HetCCL: Enabling Collective Communication For Mixed-Vendor Heterogeneous Clusters

ResearchDGX agent

arXiv:2605.31000v1 Announce Type: cross Abstract: Training Large Language Models (LLMs) on heterogeneous clusters presents significant challenges for collective communication, as hardware from multipl

IAPO: Information-Aware Policy Optimization for Token-Efficient Reasoning

SafetyDGX agent

arXiv:2602.19049v2 Announce Type: replace Abstract: Large language models increasingly rely on long chains of thought to improve accuracy, yet such gains come with substantial inference-time costs. We

If LLMs Have Human-Like Attributes, Then So Does Age of Empires II

AgentsDGX agent

arXiv:2605.31514v1 Announce Type: cross Abstract: Much research has been carried out on large language models (LLMs) and LLM-powered agentic workflows. However, many works within the field state emerg

ImmersiveTTS: Environment-Aware Text-to-Speech with Multimodal Diffusion Transformer and Domain-Specific Representation Alignment

SafetyDGX agent

arXiv:2605.30965v1 Announce Type: cross Abstract: Recent advancements in text-guided audio generation have yielded promising results in diverse domains, including sound effects, speech, and music. How

Kalimati Vegetable Price Index Forecasting with a Momentum Corrected Online Stacking Ensemble

ResearchDGX agent

arXiv:2605.30720v1 Announce Type: cross Abstract: Forecasting agricultural commodity prices in emerging economies is difficult due to high volatility, frequent supply disruptions, and strong cultural

← Previous
1…758759760761762…1018
Next →