AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
4 Aug 2026

LongCat Sparse Attention: Taming the Lightning via Streaming-aware Hierarchical Cross-Layer Indexing

Model ReleasesDGX agent

arXiv:2608.01662v1 Announce Type: cross Abstract: DeepSeek Sparse Attention (DSA) enables efficient long-context modeling through its Lightning Indexer. However, practical deployment remains constrain

LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning

Model ReleasesDGX agent

arXiv:2608.01328v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are rapidly evolving with expanded context windows and stronger reasoning capabilities, enabling multi-chart un

LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.01964v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly undertake long-horizon tasks that require sustained reasoning, tool use, and revision across many interde

Look Ahead Before You Distill: Future Trajectory Validation of Teacher Guidance for Agentic On-Policy Distillation

SafetyDGX agent

arXiv:2608.01953v1 Announce Type: new Abstract: On-policy distillation (OPD) provides teacher supervision on states visited by the student, reducing the distribution gap between training and inference

Look Up and Look Back: Hidden Attention and Latent Orientation in a Frozen Foundation Model for Panoramic SLAM

Model ReleasesDGX agent

arXiv:2608.00925v1 Announce Type: new Abstract: Monocular panoramic SLAM benefits from substantial visual overlap under large camera rotations, yet remains prone to errors caused by camera tilt, scale

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2608.02197v1 Announce Type: new Abstract: Visual representations of VLA models remain unreliable for spatially precise robotic manipulation. We uncover that vision encoders in VLAs also exhibit

Loop-Mamba: A Loop Mamba with Degradation-Aware and Shared Memory for Old Photo Restoration

Model ReleasesDGX agent

arXiv:2608.02346v1 Announce Type: new Abstract: Old photographs often suffer from multiple coupled degradations, including scratches, cracks, fading, blur, noise, and missing regions, severely degradi

LooperMuscle: Fast and Stable Learning of Humanoid Whole-Body Tracking via Structured Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2608.00820v1 Announce Type: new Abstract: FastSAC-style methods significantly reduce humanoid motion training time but often suffer from notable performance degradation compared with PPO in whol

LoopsBench: From Harness Engineering to Loop Engineering in Benchmarking Coding Agent

Model ReleasesDGX agent

arXiv:2608.00267v1 Announce Type: cross Abstract: Coding agent infrastructure is shifting from harness engineering toward loop engineering as coding agents are deployed for sustained long-horizon soft

LUT: Latent Utility Training for Visual Reasoning

SafetyDGX agent

arXiv:2608.00743v1 Announce Type: new Abstract: Multimodal large language models have advanced visual understanding, yet perception-intensive reasoning remains challenging. Recent latent visual reason

MA-HEAD-Net: Adaptive Rule-Guided Multi-Agent DRL for AoI Minimization in UAV-Assisted Emergency Networks

SafetyDGX agent

arXiv:2608.01128v1 Announce Type: cross Abstract: In post-disaster scenarios, unmanned aerial vehicles (UAVs) are critical for establishing emergency communication networks. For time-critical rescue m

Machine-Precision Prediction of Low-Dimensional Chaotic Systems from Noise-Free Data

Model ReleasesDGX agent

arXiv:2507.09652v2 Announce Type: replace-cross Abstract: Low-dimensional chaotic systems such as the Lorenz-63 model are commonly used to benchmark system-agnostic methods for learning dynamics from

Mamba Policy: Towards Efficient 3D Diffusion Policy with Hybrid Selective State Models

Model ReleasesDGX agent

arXiv:2409.07163v3 Announce Type: replace-cross Abstract: Diffusion models have been widely employed in the field of 3D manipulation due to their efficient capability to learn distributions, allowing

MANGO-Grasp: Mahalanobis Fields over Geometry-Oriented 3D Gaussians for Cross-Embodiment Dexterous Grasping

Local AiDGX agent

arXiv:2608.02014v1 Announce Type: new Abstract: Cross-embodiment dexterous grasping aims to synthesize stable grasps across heterogeneous multi-fingered hands with little or no embodiment-specific tun

Manifold-GS: Certified Hybrid Assets via Varifold-Conservative Gaussian Splatting

Model ReleasesDGX agent

arXiv:2608.00214v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) gives high-quality novel-view synthesis, but its adaptive radiance primitives are not directly usable as structured assets:

MAPLE: Metadata Augmented Private Language Evolution

Model ReleasesDGX agent

arXiv:2603.19258v2 Announce Type: replace Abstract: Differentially private (DP) fine-tuning of large language models (LLMs) requires massive compute and full model access, which rules out state-of-the

Mapping melliferous tree species in Kenya via one-class classification with hyperspectral unsupervised domain adaptation

ResearchDGX agent

arXiv:2608.02045v1 Announce Type: cross Abstract: The beekeeping sector holds significant potential for livelihood diversification among the agropastoral communities in Kenya. Melliferous tree species

MBO Scheme for Local Chan--Vese Segmentation

Local AiDGX agent

arXiv:2608.00893v1 Announce Type: new Abstract: Robust to intensity inhomogeneity, the local Chan--Vese (LCV) model extends the classical Chan--Vese (CV) image segmentation method by incorporating loc

MDTD-ArtIR: Benchmarking Image Editing and Restoration Models for Art Image Restoration under Texture-Overlay Degradations

Model ReleasesDGX agent

arXiv:2608.00736v1 Announce Type: new Abstract: Restoring severely degraded visual media still remains a formidable challenge, as existing methods often hallucinate unnatural textures and contents, st

MDWD: A Street-Level Dataset for Municipal Solid Waste Detection in Dense Urban Environments

Model ReleasesDGX agent

arXiv:2608.00257v1 Announce Type: new Abstract: Automated visual monitoring of urban environments is a growing Computer Vision research area, but municipal solid waste detection remains under-represen

Measuring in-context algorithmic reasoning in language models against an exact Bayes-optimal standard

Model ReleasesDGX agent

arXiv:2608.01575v1 Announce Type: new Abstract: Whether large language models perform genuine algorithmic reasoning or mere pattern completion is hard to test, because most benchmarks lack a ground tr

Measuring Product Quality Using Images: The CLIP Q-Score and an Application to Real Estate

ResearchDGX agent

arXiv:2608.01544v1 Announce Type: cross Abstract: The CLIP Q-score is a novel, safe, fully reproducible, and computationally efficient method for extracting objective product quality metrics from visu

MedPRESS: A Multi-turn Benchmark for Patient-Pressure-Induced Medical Sycophancy in LLMs

Model ReleasesDGX agent

arXiv:2608.02520v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for health-related advice. Existing research measures their safety with static questions rather than

MedSAM2-Anatomy: Training-Free Inference-Time Optimization for Musculoskeletal Segmentation

SafetyDGX agent

arXiv:2608.00195v1 Announce Type: cross Abstract: High-resolution 3D segmentation of hip and shoulder anatomy from CT and MRI is essential for surgical planning, yet frozen segmentation models often f

MedTextWeaver: Procedural Knowledge Evolution in Agentic Medical Text Editing

AgentsDGX agent

arXiv:2602.00740v2 Announce Type: replace Abstract: Medical text editing is essential for improving communication among diverse stakeholders in clinical settings. However, adapting LLM agents to this

MedUPS: Towards Diagnostic Assistance in Uncommon Medical Cases with Large Language Models

SafetyDGX agent

arXiv:2608.01012v1 Announce Type: new Abstract: Uncommon and off-guideline cases are difficult for clinical decision support, because physicians must make a series of management decisions under diagno

Meganeura: Portable GPU Training and Inference through Vulkan and Metal

SafetyDGX agent

arXiv:2608.01563v1 Announce Type: new Abstract: Training and deployed inference often cross export, conversion, and platform-specific runtime boundaries. Meganeura asks whether one compact native comp

MemoAct: Atkinson-Shiffrin-Inspired Hierarchical Memory-Augmented Policy for Robotic Manipulation

SafetyDGX agent

arXiv:2603.18494v2 Announce Type: replace Abstract: Memory-augmented robotic policies are essential in handling memory-dependent tasks. However, existing approaches typically rely on simply extending

MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents

AgentsDGX agent

arXiv:2608.00007v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with human-like personas is crucial for agentic applications, such as role-play and user simulation. Traditional

MemSIF: From Structured Interactions to Dual-Track Fact Memory for LLM Agents

ResearchDGX agent

arXiv:2608.01742v1 Announce Type: cross Abstract: Long-term memory is critical for LLM agents operating over long-horizon interactions. However, several persistent limitations of existing memory syste

Messages, Not Tokens: Grounded Coresets for Faithful VLM Compression

ResearchDGX agent

arXiv:2608.02134v1 Announce Type: new Abstract: Modern vision language models (VLMs) turn high-resolution images into long sequences of visual tokens. Every token traverses the language decoder and pe

MetaRoute-Bench: Evaluating Meta-Decision Policies for Agentic Workflow Routing

Model ReleasesDGX agent

arXiv:2608.00107v1 Announce Type: new Abstract: Agentic systems must repeatedly decide whether to answer directly, decompose a task, invoke a tool, execute code, delegate to a specialist, verify an in

MIDAL: Math Image Descriptions for Accessible Learning

TutorialsDGX agent

arXiv:2608.00868v1 Announce Type: new Abstract: Many open educational resources are lacking in accessibility, especially in-depth image descriptions. In subjects like Science and Mathematics, however,

MIEScore: Human-Aligned Evaluation for Multi-Source Image Editing

Model ReleasesDGX agent

arXiv:2608.02059v1 Announce Type: new Abstract: Recent advances in unified multimodal models have significantly improved text-guided image editing abilities. In particular, models such as Nano-Banana-

Mind the Gap: Zero-Query Jailbreaks via Filter-Generator Discrepancy in Text-to-Image Systems

SafetyDGX agent

arXiv:2608.00973v1 Announce Type: new Abstract: Text-to-image (T2I) systems typically have prompt-level safety filters before the generator to block unsafe requests, yet such systems remain vulnerable

MiniWorld: Democratizing the Training of Video World Models from Scratch

HardwareDGX agent

arXiv:2608.01127v1 Announce Type: new Abstract: Video world models predict future observations conditioned on historical observations and control signals, enabling long-horizon generation through auto

Minute-Scale Training for Microrobot Navigation

Model ReleasesDGX agent

arXiv:2608.00854v1 Announce Type: new Abstract: Microrobots hold significant potential for various applications, where targeted navigation is a basic requirement. Deep reinforcement learning (DRL) has

Mitigating Backdoors via Decoy Shortcuts and Knowledge Decoupling

Model ReleasesDGX agent

arXiv:2608.00732v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to deep neural networks, especially when training relies on third-party data, allowing adversaries to inject ma

Mitigating Visual Degradation in MLLMs via Spatial-Spectral Visual Anchor Learning

SafetyDGX agent

arXiv:2608.01635v1 Announce Type: new Abstract: Despite the progress of multimodal large language models (MLLMs), they continue to exhibit deficiencies in visual perception. Following visual instructi

Mitigating Visual Hallucinations in Multimodal Systems through Retrieval-Augmented Reliability-Aware Inference

SafetyDGX agent

arXiv:2606.15782v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have demonstrated strong capabilities in vision-language understanding and natural-language response

MixedComplementarityProblems.jl: A Fast, Batched, Open-Source Interior Point Solver for Mixed Complementarity Problems

Model ReleasesDGX agent

arXiv:2608.00959v1 Announce Type: cross Abstract: Mixed complementarity problems (MCPs) arise as the first-order optimality conditions of nonlinear programs and noncooperative games, and provide a nat

MMPhysVideo: Physically Plausible Video Generation Through Joint RGB-Perception Modeling

SafetyDGX agent

arXiv:2604.02817v2 Announce Type: replace Abstract: Despite advancements in generating visually stunning content, video diffusion models (VDMs) often yield physically inconsistent results due to pixel

MoCRA: Mixture of Compositional Rank-1 Atoms for 4K All-in-One Video Restoration

Model ReleasesDGX agent

arXiv:2608.01829v1 Announce Type: new Abstract: Real-world video arrives hazy, rainy, dark, or noisy, and a deployable restorer faces three demands at once: no degradation label, native 4K output, and

Model-Agnostic FDR Control via Group Gaussian Mirror and Permutation SHAP

ApplicationsDGX agent

arXiv:2608.00989v1 Announce Type: cross Abstract: Most FDR-controlled feature selection methods are designed for coordinate-wise hypotheses, where each feature has a single weight or importance score.

Modeling Unknown Nonlocal PDE Systems via Flow Map Learning

TutorialsDGX agent

arXiv:2608.00400v1 Announce Type: new Abstract: Nonlocal partial differential equations arise in many applications but are often difficult to model and learn because of the presence of nonlocal operat

Models as Tools: An Agentic Coordination Framework for Unified Multimodal Visual Tracking

Model ReleasesDGX agent

arXiv:2608.00847v1 Announce Type: new Abstract: Most current visual trackers adopt a matching-based architecture trained exclusively on tracking datasets, whose performance gains depend heavily on the

MonitorVLM-v2: A Deployed Vision-Language Framework for Real-Time Safety Violation Detection

SafetyDGX agent

arXiv:2608.00975v1 Announce Type: new Abstract: Large vision--language models (VLMs) can reason step by step about complex visual scenes, but this open-ended, autoregressive chain-of-thought (CoT) app

MoRAL: Sensor-Grounded BEV Reasoning for Compact VLMs toward Edge-Oriented Autonomous Driving

Model ReleasesDGX agent

arXiv:2608.02449v1 Announce Type: new Abstract: Deploying vision-language models (VLMs) for safety-critical spatial reasoning on resource-constrained autonomous driving platforms requires both compact

Morphology Aware Reversible Semantic Tokenization and Hierarchical Word Composition for Tamil Language Models

Model ReleasesDGX agent

arXiv:2608.01153v1 Announce Type: new Abstract: Statistical subword tokenizers can process arbitrary text, but their units need not align with lexical or grammatical structure. This is especially impo

Motion Beyond Morphology: Bootstrapping Cross-Category Motion Transfer from Abstract Motion Representations

ResearchDGX agent

arXiv:2608.01628v1 Announce Type: new Abstract: Video motion transfer aims to animate a target object using dynamics from a reference video. Existing formulations largely rely on fixed structural corr

Motion Planning for Mobile Manipulators Navigating Doorways via Model Predictive Control

ResearchDGX agent

arXiv:2608.00206v1 Announce Type: new Abstract: Navigating doorways is a fundamental capability for mobile manipulators operating in human environments, requiring coordinated motion between the mobile

Move What Matters: Parameter-Efficient Domain Adaptation via Optimal Transport Flow for Collaborative Perception

Model ReleasesDGX agent

arXiv:2602.11565v5 Announce Type: replace Abstract: Efficient domain adaptation remains a fundamental challenge for deploying multi-agent systems across diverse environments in Vehicle-to-Everything (

Multi-Source Dynamic Graph Learning for Compound-Flood Forecasting in Managed Coastal Systems

SafetyDGX agent

arXiv:2608.01775v1 Announce Type: new Abstract: Compound flooding in managed coastal systems is influenced by hydrological conditions and water-management activity observed across multiple monitoring

Multi-View Unified Camera Fields: Geometry-Shaped Action-Facing Representations for RGB-Only Multi-Camera VLA Policies

ApplicationsDGX agent

arXiv:2608.01826v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong generalization in robotic manipulation, yet complex contact-rich tasks often benefit from multi-ca

MUSS: Multilevel Subset Selection for Relevance and Diversity

ApplicationsDGX agent

arXiv:2503.11126v4 Announce Type: replace Abstract: The problem of relevant and diverse subset selection has a wide range of applications, including recommender systems and retrieval-augmented generat

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

SafetyDGX agent

arXiv:2603.14686v2 Announce Type: replace Abstract: Human-Object Interaction (HOI) video reenactment aims to transfer the interaction dynamics of a source video to a novel target object while preservi

Native Multilingual Chain-of-Thought Reasoning in Low-Resource Southeast Asian Languages

SafetyDGX agent

arXiv:2608.00533v1 Announce Type: new Abstract: Large Language Models have achieved substantial progress in reasoning capabilities. Yet in low-resource native settings, many suffer from cross-lingual

Near-Optimal Reinforcement Learning for Constrained Recurrence Objectives

SafetyDGX agent

arXiv:2511.19849v2 Announce Type: replace-cross Abstract: Recurrence objectives, where a target region must be visited infinitely often, are a fundamental class of specifications for Markov decision p

NetDiff: Graph Diffusion with Improved Global Capabilities to Generate and Update Mobile Network Topologies

ResearchDGX agent

arXiv:2410.08238v2 Announce Type: replace-cross Abstract: We introduce NetDiff, a node-conditioned denoising diffusion model that generates directional link topologies and a two-slot transmit/receive

Network Information Enhances Unreliable News Domain Detection

ResearchDGX agent

arXiv:2608.02399v1 Announce Type: cross Abstract: Content-based detection of unreliable news is increasingly difficult, as low-reliability sources mimic credible journalism and generative AI makes fab

← Previous
1…979899100101…998
Next →