AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
5 Aug 2026

ToolLIFT: Lifting Tool-Specific Trajectories into Function-Level Graphs for Generalizable Tool Planning

AgentsDGX agent

arXiv:2608.03468v1 Announce Type: new Abstract: Historical tool-use trajectories provide valuable experience for large language model (LLM) agents to plan and coordinate tool usage. Existing approache

Topological Simplification in Predictive Coding Networks

ResearchDGX agent

arXiv:2608.02816v1 Announce Type: new Abstract: We study the topology of learned representations in predictive coding networks (PCNs), a neuro-inspired bidirectional architecture, using a quantitative

Towards Robust Tool Use in Agents via Experience-Driven Adaptive Guidance

AgentsDGX agent

arXiv:2608.03403v1 Announce Type: new Abstract: The performance bottleneck of agents is increasingly shifting from model capability to the robustness of their execution processes. Tools play a central

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Unexpected Crash-Looping of Multi-GPU Box: Caught by Hand-Scrutinized Logs - A Tale of Overridden Keep-Alive Policy and Eviction Thrashing

Local AiDGX agent

I've been freelancing for over a decade now, and I can't stress enough the importance of thorough investigation when dealing with strange software behaviors. Recently, I ran into an issue where my mul

Unified Visuomotor Targets: Supervising VLAs Beyond Physical Actions

SafetyDGX agent

arXiv:2608.03563v1 Announce Type: new Abstract: VLA models are trained to predict robot actions from visual and language observations. This is a natural choice, but it creates a mismatch: VLMs encode

UNVaMP: Neural Knowledge Tracing with Variational Regularization of Latent Knowledge Dynamics

ApplicationsDGX agent

arXiv:2608.03811v1 Announce Type: new Abstract: We introduce the Unified Neural Variational Measurement of Proficiency (UNVaMP) architecture, a knowledge tracing method that integrates observed studen

Variational Approximated Restricted Maximum Likelihood Estimation for Spatial Data

ResearchDGX agent

arXiv:2604.07635v2 Announce Type: replace-cross Abstract: This research considers a scalable inference for spatial data modeled through Gaussian intrinsic conditional autoregressive (ICAR) structures.

Very interesting to see @JeffDean's pitch deck. Just look at those open science and engineering problems. Lots to advance there with automat…

ResearchDGX agent

Very interesting to see @JeffDean's pitch deck. Just look at those open science and engineering problems. Lots to advance there with automated ML engineering. AI for science and engineering is just ge

What Language Does and What the Evidence Supports: A Functional Role Taxonomy and Evidence Audit of Language Grounding in Embodied Agents

ResearchDGX agent

arXiv:2608.03099v1 Announce Type: new Abstract: Foundation models place language throughout embodied agents, but its presence does not show what it contributes or how well that contribution is grounde

When and Where to Look: Adaptive Visual Evidence Scheduling for Efficient Long Video Understanding

Local AiDGX agent

arXiv:2608.03918v1 Announce Type: cross Abstract: Efficient long-video understanding requires vision--language models (VLMs) to reason over a small number of frames selected as sparse visual evidence.

When Compression Scores Cannot Decide: Information Boundaries for Group-Robust LLM Pruning

Local AiDGX agent

arXiv:2608.02940v1 Announce Type: new Abstract: A reproducible compression statistic can still select the wrong candidate. A dense pruning score with 0.906 split-half reliability predicted a 16.1% gai

When Oracle Conditioning Misleads Deployment: Conditioning-Availability Bias in Echocardiographic Segmentation

SafetyDGX agent

arXiv:2608.03342v1 Announce Type: cross Abstract: Conditional segmentation models may be trained and evaluated with auxiliary signals cleaner than those available at deployment. We study this protocol

When Teachers Mislead: Spurious-Signal-Aware On-Policy Distillation

SafetyDGX agent

arXiv:2608.03632v1 Announce Type: new Abstract: On-Policy distillation (OPD) transfers teacher capabilities by supervising student-sampled trajectories with dense token-level teacher signals. Recent s

4 Aug 2026

3D-CovDiffusion: 3D-Aware Diffusion Policy for Coverage Path Planning

SafetyDGX agent

arXiv:2510.03011v2 Announce Type: replace Abstract: Diffusion models have shown strong potential for robot skill learning, yet their role in coverage path planning remains underexplored. In industrial

3DZip: Spatial-Aware Feature Diversity-Guided Token Compression for 3D Question Answering

ResearchDGX agent

arXiv:2608.01185v1 Announce Type: new Abstract: Recent 3D vision-language models (3D VLMs) construct geometry aware tokens by projecting 2D visual features into world coordinates, enabling spatial rea

Abstention as an Action Can Kill Both the Reward Gradient and the KL Anchor: Collapse Law and Repair for Error-Penalized Reinforcement Learning

AgentsDGX agent

arXiv:2608.00301v1 Announce Type: cross Abstract: Error-penalized scoring rules (+1 for a correct answer, -lambda for a wrong one, 0 for abstaining) are increasingly prescribed against hallucination:

Assessing the Impacts of Imperfect Datasets on Client Selections in Federated Learning

Local AiDGX agent

arXiv:2608.02250v1 Announce Type: new Abstract: Federated learning (FL) is a popular distributed learning framework where multiple clients perform local training and a server aggregates the locally up

Attend to Your Own Thoughts: Breaking the Barrier for Post-Training Quantization of Reasoning LLMs through the Lens of 1.58-Bit Quantization

ResearchDGX agent

arXiv:2608.01078v1 Announce Type: new Abstract: We propose ScaleQ-1.58, a scalable ternary post-training quantization (PTQ) framework for reasoning LLMs. Its core insight stems from an empirical findi

Benign Overfitting in Linear Classifiers with a Bias Term

SafetyDGX agent

arXiv:2511.12840v2 Announce Type: replace-cross Abstract: Overparameterized models often generalize well even when they interpolate noisy training data. This is known as benign overfitting. For linear

Breaking the Statistical Similarity Trap in Extreme Convection Detection

SafetyDGX agent

arXiv:2509.09195v2 Announce Type: replace-cross Abstract: Current evaluation metrics for deep learning weather models create a 'Statistical Similarity Trap', rewarding blurry predictions while missing

Can AI Agents Simulate A/B Test Outcomes? A Validation Framework for Agentic Experimentation

AgentsDGX agent

arXiv:2608.02345v1 Announce Type: new Abstract: A/B testing remains the standard for rolling out new features in the technology industry. Each experiment, however, consumes real traffic, engineering e

CascadeLUT: Information-Ordered Streaming Inference for Bandwidth-Constrained FPGAs

Local AiDGX agent

arXiv:2608.00720v1 Announce Type: cross Abstract: Mapping neural networks to FPGAs enables low-latency, energy-efficient inference, particularly for lookup table (LUT)-based models that eliminate mult

Coverage-Driven Adaptive Keyframe Selection for Video Understanding

TutorialsDGX agent

arXiv:2608.00714v1 Announce Type: new Abstract: Recent advances in large vision-language models (LVLMs) have enabled long-video understanding and analysis. However, processing the large number of fram

CoWAM: Coordination Contracts for Selective Policy Intervention with WAMs

SafetyDGX agent

arXiv:2608.02578v1 Announce Type: cross Abstract: World Action Models (WAMs) augment robot policies with action-conditioned predicted futures, but a plausible future alone does not justify changing th

Credit the Right Box: Marginal Contribution Assignment for Structured Visual Perception

ResearchDGX agent

arXiv:2608.01055v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are increasingly expected to solve structured perception tasks that require visual recognition, language-to-obj

Cultural Awareness is Represented but Not Decoded: Tracing Mythological Knowledge across 18 Open-Source LLMs

ResearchDGX agent

arXiv:2608.02486v1 Announce Type: new Abstract: Open-source LLMs reliably name Zeus, Jupiter, and Thor, but recover their counterparts in less-represented traditions like Finnish, Slavic, Egyptian, or

D^2-4DGS: Dual-Depth Guided Sparse-Camera 4D Gaussian Splatting

ResearchDGX agent

arXiv:2608.01588v1 Announce Type: new Abstract: Dynamic 4D Gaussian Splatting has emerged as an efficient representation for dynamic novel view synthesis through explicit scene modeling and real-time

Decisions over Sequences: Computability and Choice

ResearchDGX agent

arXiv:2203.00070v3 Announce Type: replace-cross Abstract: We develop a framework to study situations where decision makers face alternatives sequentially. Within this framework, we focus on endogenous

Deep Research Pretraining via Predictive Navigation

SafetyDGX agent

arXiv:2608.00432v1 Announce Type: new Abstract: Deep research agents are often trained on expensive, environment-grounded tool-use trajectories that require repeated retrieval, document inspection, an

DeepDefense: Robust Learning via Layer-Wise Gradient-Feature Alignment

SafetyDGX agent

arXiv:2511.13749v2 Announce Type: replace Abstract: Deep neural networks are known to be vulnerable to adversarial perturbations, which are small, carefully crafted inputs that lead to incorrect predi

Demystifying When and Why VLAs Fail in Contact-Rich Tasks and How to Fix Them

SafetyDGX agent

arXiv:2608.01402v1 Announce Type: new Abstract: We address the problem of understanding when and why Vision-Language-Action models struggle with contact-rich manipulation tasks that require precise ph

Detecting Nonproperness of Likelihood Equations

ResearchDGX agent

arXiv:2608.01976v1 Announce Type: cross Abstract: Given an algebraic statistical model, a challenging problem is classifying the data according to the number of positive critical points of the likelih

DiffuseAgent-MI: Distributionally-Grounded,Tool-Integrated Self-Evolving Agents for Faithful Visual Reasoning

SafetyDGX agent

arXiv:2608.00540v1 Announce Type: new Abstract: Tool-integrated vision-language agents have made remarkable progress on compositional and multi-step visual reasoning. Yet their outputs frequently exhi

Diffusion-Based Body Schema Learning Enabling Abnormal-State Adaptation in Musculoskeletal Robots

TutorialsDGX agent

arXiv:2608.01029v1 Announce Type: new Abstract: Musculoskeletal robots require an internal body schema that remains consistent under a wide range of physical state changes, including abnormalities suc

Disentangled Contrastive Learning for Zero-Shot Multilingual Dense Retrieval

SafetyDGX agent

arXiv:2608.02189v1 Announce Type: cross Abstract: Multilingual dense retrieval aims to handle queries and documents across different languages based on a unified retriever model. The challenge lies in

Distributional Matching for Vector Quantization: A Unified Theoretical and Empirical Framework

ResearchDGX agent

arXiv:2607.15933v2 Announce Type: replace Abstract: The effectiveness of modern visual representation learning and autoregressive models critically depends on vector quantization (VQ), which discretiz

Does the Competitive Component of Adversarial Self-Play Improve Legal Reasoning? A Controlled Negative Result

ApplicationsDGX agent

arXiv:2608.01559v1 Announce Type: cross Abstract: Adversarial self-play is an appealing recipe for legal reasoning: have a student model draft an argument, have an adversary attack it, and reward the

DreamTraj: Generating 6-DoF Object Trajectories by Reading Unrendered Video Diffusion Latents

ResearchDGX agent

arXiv:2608.00486v1 Announce Type: new Abstract: Accurate prediction of object trajectories during manipulation is essential for closing the perception-action loop. Progress is limited on two fronts: a

Dynamic Resolution Routing for Efficient Egocentric Grounding

Local AiDGX agent

arXiv:2608.01638v1 Announce Type: new Abstract: Egocentric visual grounding requires high-resolution inputs to localize small objects. However, scaling Multimodal Large Language Models to this domain

Exploiting Intrinsic Duality for Multi-Hop Question Generation

SafetyDGX agent

arXiv:2608.00712v1 Announce Type: new Abstract: Multi hop question generation (MQG) aims to generate questions from multiple given documents and target answers, whereas question answering (QA) focuses

FactorJEPA: Factorizing Monolithic Futures into Layout-Agent-Interaction Channels for Crowded and Chaotic Global South Urban Worlds

AgentsDGX agent

arXiv:2608.01049v1 Announce Type: cross Abstract: World models have attracted significant attention for their ability to capture and predict the structure and dynamics of the physical world. In this e

FAST-GS: Frequency Aware Space-time Gaussian Splatting for Photorealistic Dynamic Novel View Synthesis

Local AiDGX agent

arXiv:2608.01958v1 Announce Type: new Abstract: 4D Gaussian Splatting (4DGS) excels in dynamic 3D reconstruction and real-time novel view synthesis via efficient 4D Gaussian representations and parall

FAU at ImageCLEF 2026 Task on Multimodal Reasoning Robust Candidate Scoring and Concise Multilingual Visual Answering

ResearchDGX agent

arXiv:2608.01664v1 Announce Type: new Abstract: We present our ImageCLEF 2026 Multimodal Reasoning system for the Visual Multiple Choice Question Answering (Visual MCQ) and Visual Open Question Answer

From Chains to Trees: Parent-Conditioned Drafting for Semi-Autoregressive Speculative Decoding

ResearchDGX agent

arXiv:2608.02123v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference only when drafted continuations survive target-model verification. Semi-autoregressive drafters such as D

From fragmented data to actionable design: Physics-calibrated learning for plastic upcycling

ResearchDGX agent

arXiv:2608.02402v1 Announce Type: new Abstract: Thermochemical upgrading of plastic waste is a key upcycling pathway, yet the experimental literature is fragmented by heterogeneous conditions and inco

From Pixels to PCells: A Neurosymbolic Approach to Photonic Component Creation

ResearchDGX agent

arXiv:2608.00084v1 Announce Type: new Abstract: We present PixCell, a neurosymbolic system in which multimodal agents convert a visually presented photonic component into a parametric program over a s

GAPSL: A Gradient-Aligned Parallel Split Learning over Data-Heterogeneous Edge Computing Systems

SafetyDGX agent

arXiv:2603.18540v2 Announce Type: replace Abstract: The increasing complexity of neural networks poses significant challenges for democratizing federated learning (FL) on resource-constrained edge dev

Generative Brownian Bridge Diffusion In Motion Space For Enhanced Myocardial Strain Analysis

TutorialsDGX agent

arXiv:2608.01677v1 Announce Type: new Abstract: Myocardial strain analysis of cardiac magnetic resonance (CMR) images provides an important tool for evaluating cardiac function. However, current techn

Gimbal360: Canonicalizing Planar Diffusion for Spherical Panorama Completion

ResearchDGX agent

arXiv:2603.23179v2 Announce Type: replace Abstract: Diffusion models provide powerful priors for 2D image completion, but these priors are learned on bounded planar images and do not transfer directly

Grounded Vision-Language Interpreter for Long-Horizon Bimanual Task and Motion Planning

SafetyDGX agent

arXiv:2506.03270v3 Announce Type: replace Abstract: While recent advances in vision-language models have accelerated language-guided robot planning, their black-box nature lacks the safety guarantees

HiResNets: Native Full-HD Video Recognition with Foveal Residual Streams

TutorialsDGX agent

arXiv:2608.02140v1 Announce Type: new Abstract: Much of the recent progress in image and video recognition has come at the cost of memory: larger models, increased resolution, and longer temporal cont

Hybrid Quantum CNN for Cross-Sensor Spaceborne Volcanic Thermal Activity Recognition Worldwide

ResearchDGX agent

arXiv:2608.00069v1 Announce Type: cross Abstract: As Earth Observation (EO) enters the Big Data era, the exponential volume of daily satellite imagery poses significant computational and storage chall

InstancePin: Instance-Addressable Layout-to-Image Diffusion via Coordinate Pinning

ResearchDGX agent

arXiv:2608.00588v1 Announce Type: new Abstract: Layout-to-image diffusion models have achieved impressive semantic controllability by conditioning generation on category-level segmentation maps. Howev

Latent-Regime Bias Auditing for Volatility Forecasting

SafetyDGX agent

arXiv:2608.01599v1 Announce Type: new Abstract: Volatility forecasts are commonly evaluated with aggregate accuracy metrics such as RMSE and MAE, but these metrics can hide conditional failures that m

Latent Softmax for Data-Efficient Phoneme-Based Multilingual ASR Across Tonal and Non-Tonal Languages

ResearchDGX agent

arXiv:2608.01281v1 Announce Type: cross Abstract: Phoneme-based multilingual automatic speech recognition (ASR) can share acoustic evidence across languages more directly than language-specific subwor

LEAP: Lean Environment-Feedback via Adaptive Pruning for Code RL in GPU Kernel Generation

SafetyDGX agent

arXiv:2608.01804v1 Announce Type: new Abstract: Post-training large language models (LLMs) via reinforcement learning (RL) has significantly advanced code generation capabilities. To bypass the heavy

Learning-Based Collaborative MEC for LLM Inference with Soft-Deadline Awareness via Transformer-Enhanced PPO

SafetyDGX agent

arXiv:2608.02031v1 Announce Type: cross Abstract: This paper investigates collaborative mobile edge computing (MEC) servers for large language model (LLM) inference under soft deadline constraints. In

Learning to Persuade Privately Informed Receivers

ResearchDGX agent

arXiv:2607.28342v1 Announce Type: cross Abstract: Bayesian persuasion studies how an informed sender can influence the behavior of a receiver through strategic information disclosure. Standard models

Leveraging Synthetic Data for Question Answering with Multilingual LLMs in the Agricultural Domain

ResearchDGX agent

arXiv:2507.16974v3 Announce Type: replace Abstract: Enabling farmers to access accurate agriculture-related information in their native languages in a timely manner is crucial for the success of the a

LLM generation novelty through the lens of semantic similarity

ResearchDGX agent

arXiv:2510.27313v3 Announce Type: replace-cross Abstract: Generation novelty is a key indicator of an LLM's ability to generalize, yet measuring it against full pretraining corpora is computationally

← Previous
1…736737738739740…1034
Next →