AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
Human
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
7 Aug 2026

MM-ISTS: Cooperating Irregularly Sampled Time Series Forecasting with Multimodal Vision-Text LLMs

SafetyDGX agent

arXiv:2603.05997v2 Announce Type: replace-cross Abstract: Irregularly sampled time series (ISTS) are widespread in real-world scenarios, exhibiting asynchronous observations on uneven time intervals a

MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation

ApplicationsDGX agent

arXiv:2603.25406v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models map visual observations and natural-language instructions to robot actions; however, hierarchical and autoregres

MoCA: Implicit Social Context Analysis

Model ReleasesDGX agent

arXiv:2608.05825v1 Announce Type: new Abstract: Human social communication, such as affection and intent, is often conveyed in highly implicit ways, where underlying meanings are expressed through ind

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Mood Matters: How Syntactic Sensitivity Undermines Safety Alignment

SafetyDGX agent

arXiv:2608.05409v1 Announce Type: new Abstract: Large language models typically undergo post-training to align them with safety policies but there exist many sophisticated jailbreaks that sidestep est

MOSAIK: Multi-Patch Content-Aware Spatial Allocation of Image Tokens for Efficient Generation

ResearchDGX agent

arXiv:2608.05450v1 Announce Type: new Abstract: Pixel-space diffusion models avoid the reconstruction ceiling of latent diffusion models by generating directly in image space. However, their substanti

MS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classification

Model ReleasesDGX agent

arXiv:2608.05196v1 Announce Type: new Abstract: Multiple sclerosis (MS) is diagnosed through clinical assessment, magnetic resonance imaging, laboratory evidence when appropriate, and exclusion of bet

Multi-Agent Reinforcement Learning for Online Traffic Scheduling in Time-Sensitive Application

SafetyDGX agent

arXiv:2608.05346v1 Announce Type: cross Abstract: Time-sensitive networking (TSN) is increasingly integrated into mobile edge computing (MEC) to support applications with stringent latency requirement

Multi-Agent Transformer for Queue-Level XR Traffic Scheduling in TSN Networks

AgentsDGX agent

arXiv:2608.05340v1 Announce Type: cross Abstract: Time-Sensitive Networking (TSN) and Mobile Edge Computing (MEC) hold strong potential for enabling ultra-reliable low-latency communication for time-s

Multi-Representation Geometric Hierarchy Fusion: An Implicit-Submap Driven Framework for Resilient 3D Place Recognition

Model ReleasesDGX agent

arXiv:2506.14243v4 Announce Type: replace Abstract: LiDAR-based place recognition is critical for long-term autonomous driving without GPS. Existing handcrafted feature methods face dual limitations.

Multi-Year Geospatial Reasoning using Interannually-Consistent Historical Predictions as a Free Input Modality

ResearchDGX agent

arXiv:2608.05979v1 Announce Type: new Abstract: Machine learning, and deep networks in particular, are increasingly used to derive higher-level Earth observation (EO) products such as annual land-cove

Multivariate Time Series Forecasting needs Cross Variable Loss

ResearchDGX agent

arXiv:2608.05742v1 Announce Type: cross Abstract: Multivariate time series forecasting presents unique challenges because future variables often co-evolve under shared system dynamics. While existing

Muon on the Stiefel Manifold Admits an Exact Closed-Form Update

ResearchDGX agent

arXiv:2608.06218v1 Announce Type: cross Abstract: We study Muon, a recently proposed matrix-aware optimization method, in the context of the Stiefel manifold. This manifold consists of matrices with o

NavTrust: Benchmarking Trustworthiness for Embodied Navigation

Model ReleasesDGX agent

arXiv:2603.19229v2 Announce Type: replace-cross Abstract: There are two major categories of embodied navigation: Vision-Language Navigation (VLN), where agents navigate by following natural language i

Near-sensor Computing for Rapid Visuotactile Perception

ResearchDGX agent

arXiv:2608.05725v1 Announce Type: new Abstract: Visuotactile sensors reconstruct dense contact geometry from measured surface gradients, but host-based processing increases power consumption and intro

Negotiating Risk Boundaries in AI for Policing Through Mixed-Stakeholder Deliberation

SafetyDGX agent

arXiv:2608.05418v1 Announce Type: new Abstract: AI tools are being increasingly adopted in policing in the UK and worldwide. Racial bias is a known and well-documented risk, yet representatives of aff

NeSy-RAG: Neuro-Symbolic RAG for Explainable Question Answering

Model ReleasesDGX agent

arXiv:2608.06292v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves question answering by grounding large language models (LLMs) in external knowledge such as text corpora. H

Neuro-Symbolic Closed-Loop Control of Laser Powder Bed Fusion with an In-Loop Ontology

Model ReleasesDGX agent

arXiv:2608.05773v1 Announce Type: new Abstract: A geometry-conditioned, neuro-symbolic closed-loop architecture is proposed for laser powder bed fusion, in which a standards-aligned ontology operates

NeuroAdaptTrainer: A Fiji/ImageJ Plugin for YOLO-Based Neuron Segmentation, InteractiveCorrection and Transfer Learning

ResearchDGX agent

arXiv:2608.05226v1 Announce Type: new Abstract: Neuron counting and segmentation in microscopy images of neuronal cultures is a routine and time-consuming task in neuroscience research, traditionally

nnMIL: A generalizable multiple instance learning framework for computational pathology

ApplicationsDGX agent

arXiv:2511.14907v2 Announce Type: replace Abstract: Computational pathology holds substantial promise for improving diagnosis and guiding treatment decisions. Recent pathology foundation models enable

Nonvisual Classification of Ground-Condition by Artificial Proprioception in an Amoeba-Inspired Autonomous Walking Robot

AgentsDGX agent

arXiv:2608.05684v1 Announce Type: cross Abstract: Nonvisual classification of ground condition based on a multimodal sensing approach was investigated for an amoeba-inspired autonomous walking robot.

Observation-Grounded Self-Predictive Reinforcement Learning for Visual Continuous Control

SafetyDGX agent

arXiv:2608.05989v1 Announce Type: new Abstract: Sample-efficient policy learning from pixels is a long-standing challenge in reinforcement learning (RL). Recent dynamics-based representation learning

omega-0: A Latent Predictive World Action Model for Concurrent Humanoid Loco-Manipulation

ApplicationsDGX agent

arXiv:2608.06375v1 Announce Type: new Abstract: Humanoid household tasks often require concurrent loco-manipulation, where the robot must move, adjust posture, maintain balance, and manipulate objects

OmniMech: All-in-one Multimodal Mechanical Benchmark for 3D Reconstruction

Model ReleasesDGX agent

arXiv:2608.05539v1 Announce Type: new Abstract: Recent vision-language models (VLMs) can generate executable CAD programs from images, but existing methods mainly target coarse, general-purpose 3D obj

On-Policy Delta Distillation for Multilingual Math Reasoning

SafetyDGX agent

arXiv:2608.05802v1 Announce Type: new Abstract: On-Policy Distillation (OPD) is emerging as a promising alternative to reinforcement learning for LLM post-training, yet its effectiveness in multilingu

On-Policy Self-Distillation without Any Supervision

SafetyDGX agent

arXiv:2608.06296v1 Announce Type: new Abstract: On-policy (Self-)Distillation (OPD / OPSD) has shown strong potential for post-training large language models (LLMs). However, existing methods still re

On Same-Sample and Independent-Sample Stochastic Extragradient for Monotone Variational Inequalities

ResearchDGX agent

arXiv:2608.06182v1 Announce Type: cross Abstract: We study stochastic extragradient (SEG) methods for solving monotone variational inequality problems (VIPs) over a feasible set. Although extragradien

On the Anisotropy of Score-Based Generative Models

ResearchDGX agent

arXiv:2510.22899v2 Announce Type: replace Abstract: We investigate the role of network architecture in shaping the inductive biases of modern score-based generative models. To this end, we introduce t

Once a Response, Always a Response: Detecting LLM-generated Text via Latent Prompt Restoration

ResearchDGX agent

arXiv:2608.05741v1 Announce Type: cross Abstract: Large language models (LLMs) can generate fluent and convincing text at scale, creating growing risks for misinformation dissemination, educational mi

One Leak Away: How Pretrained Model Exposure Amplifies Jailbreak Risks in Finetuned LLMs

Model ReleasesDGX agent

arXiv:2512.14751v3 Announce Type: replace-cross Abstract: Finetuning pretrained large language models (LLMs) has become the standard paradigm for developing downstream applications. However, its secur

One Qubit Can Beat One Bit: Quantum Advantage for Post-Training Quantization

ResearchDGX agent

arXiv:2608.05240v1 Announce Type: cross Abstract: One-bit post-training quantization represents each weight using only its sign, requiring all deployment contexts to share the same binary weight matri

One Ranking, Any Budget: Matryoshka Evidence-to-Context Frame Selection for Long-Video Understanding

ResearchDGX agent

arXiv:2608.05707v1 Announce Type: new Abstract: Frame selection is essential for applying Large Multimodal Models (LMMs) to long videos due to severe frame redundancy and limited context windows. Sinc

Online Reasoning Calibration: Test-Time Training Enables Generalizable Conformal LLM Reasoning

ResearchDGX agent

arXiv:2604.01170v2 Announce Type: replace-cross Abstract: While test-time scaling has enabled large language models to solve highly difficult tasks, state-of-the-art results come at exorbitant compute

OPERA: Operator-residual feedback for reliable autonomous optical experiments with language-model agents

AgentsDGX agent

arXiv:2608.05990v1 Announce Type: new Abstract: Autonomous agents choose actions using scores that may not reflect experimental success. We developed OPERA, an operator-residual framework for optical

Operating Multi-Node Full Fine-Tuning on NVIDIA B300: A Field Report on Telemetry-Based Triage, Negative Results, and Operational Hardening

Model ReleasesDGX agent

arXiv:2608.05944v1 Announce Type: cross Abstract: We report operational experience full-fine-tuning a 32.76B-parameter dense model (Qwen3-32B) on 16 x NVIDIA B300 (two nodes, FSDP / ZeRO-3) -- among t

Optimal or Greedy Decision Trees? Revisiting their Objectives, Tuning, and Performance

ResearchDGX agent

arXiv:2409.12788v3 Announce Type: replace Abstract: Recently there has been a surge of interest in optimal decision tree (ODT) methods that globally optimize accuracy directly, in contrast to traditio

Optimal Rates for Learning with Monotone Adversaries

ResearchDGX agent

arXiv:2608.06337v1 Announce Type: cross Abstract: A monotone adversary observes an i.i.d. labeled sample and appends a finite number of further examples of its choice, every one of them labeled correc

OrchestraBench: Evaluating Multi-Agent Orchestration Failure Modes, Recovery, and Decomposition Quality

Model ReleasesDGX agent

arXiv:2608.05263v1 Announce Type: new Abstract: Multi-agent orchestration frameworks are moving from demos to production, yet benchmarks typically report task accuracy without diagnosing why a pipelin

Ordered Diffusion for 3D Human Registration

SafetyDGX agent

arXiv:2608.05804v1 Announce Type: new Abstract: 3D human registration has historically been treated as a regression task, assuming a unique ground-truth alignment exists between the template and an in

OTLesMix: Wasserstein Barycenter and Optimal Transport Map for Synthetic Lesion Generation with Diverse Shapes and Locations

ResearchDGX agent

arXiv:2608.06264v1 Announce Type: new Abstract: The development of deep learning over the past decade has revolutionized medical imaging segmentation, allowing the extraction of precise descriptors fr

Otter: A Time-Aware, History-Conditioned Human Chess AI

Model ReleasesDGX agent

arXiv:2608.05206v1 Announce Type: new Abstract: Otter is a 15.3M-parameter human chess AI that predicts human move selection by modeling play as a time-aware, sequential process rather than treating e

Overcoming Attention Drift: Homogeneity-Heterogeneity Guided Feature Aggregation for Low-Light Remote Sensing Image Enhancement

SafetyDGX agent

arXiv:2608.05843v1 Announce Type: new Abstract: Restoring high-fidelity remote sensing imagery from extreme low-light degradation is indispensable for reliable Earth observation and downstream machine

PaCoNet: Deep Data Extraction for Parallel Coordinates

ResearchDGX agent

arXiv:2608.06030v1 Announce Type: new Abstract: Extracting data from visualizations has long challenged computer vision, with current research focused on bar, line, and pie charts, among other low-dim

PaDoc: Layout-Grounded Parallel Decoding for Document Parsing

HardwareDGX agent

arXiv:2608.06146v1 Announce Type: new Abstract: End-to-end document parsers provide a unified interface, but serialize page layouts and regional contents into one autoregressive sequence. This formula

Parallel-in-Time Nonlinear Optimal Control via GPU-native Sequential Convex Programming

HardwareDGX agent

arXiv:2603.10711v3 Announce Type: replace Abstract: Real-time solution of nonlinear optimal control problems remains challenging on embedded robotic hardware, where conventional solvers often rely on

Parameter-Efficient Semantic Augmentation for Enhancing Open-Vocabulary Object Detection

Model ReleasesDGX agent

arXiv:2604.04444v2 Announce Type: replace Abstract: Open-vocabulary object detection (OVOD) enables models to detect any object category, including unseen ones. Benefiting from large-scale pre-trainin

Path Planning of Cleaning Robot with Reinforcement Learning

SafetyDGX agent

arXiv:2208.08211v2 Announce Type: replace-cross Abstract: Recently, as the demand for cleaning robots has steadily increased, therefore household electricity consumption is also increasing. To solve t

PathCover: A Fast Convex Decomposition along a Path via Randomized Iterative Space Partitioning (RISP) on Point Clouds

AgentsDGX agent

arXiv:2608.05586v1 Announce Type: new Abstract: Autonomous robot navigation requires the rapid generation of obstacle-free regions for trajectory planning. However, existing corridor generators strugg

Patient Pose Assessment Using a CT-Based Framework for Synthetic Data Generation

ResearchDGX agent

arXiv:2608.06126v1 Announce Type: new Abstract: An adequate diagnostic quality of radiographs is essential for reliable diagnoses and treatment planning. The patient's pose during radiography is one o

PD-GS: Phoneme-Driven 3DGS for Audio-Driven Talking Heads

SafetyDGX agent

arXiv:2608.05218v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) enables fast, photorealistic talking-head rendering, yet accurate lip articulation remains elusive: mouth motion is often o

Perfect reconstruction of sparse signals using nonconvexity control and one-step RSB message passing

Model ReleasesDGX agent

arXiv:2512.17426v2 Announce Type: replace-cross Abstract: We consider sparse signal reconstruction via minimization of the smoothly clipped absolute deviation (SCAD) penalty, and develop one-step repl

Persona-Pruner: Sculpting Lightweight Models for Role-Playing

ApplicationsDGX agent

arXiv:2606.14695v2 Announce Type: replace-cross Abstract: Language Models (LMs) have shown remarkable potential as role-playing chatbots, delivering consistent, stylized interactions when given a spec

Personalized Deep Research Query Refinement with Graph-Scaffolded Evidence Grounding

SafetyDGX agent

arXiv:2608.05876v1 Announce Type: new Abstract: User requests serve as research specifications for deep research agents, shaping what evidence to seek and how to synthesize it. In personalized deep re

Perturbation Sensitivity at Convergence: A Simple Signal for Identifying Spuriously Correlated Samples

ResearchDGX agent

arXiv:2608.05419v1 Announce Type: cross Abstract: Models trained by empirical risk minimization on data containing spurious correlations achieve high average accuracy while failing on subpopulations w

PhaseCoder: Microphone Geometry-Agnostic Spatial Audio Understanding for Multimodal LLMs

Model ReleasesDGX agent

arXiv:2601.21124v2 Announce Type: replace-cross Abstract: Current multimodal LLMs process audio as a mono stream, ignoring the rich spatial information essential for embodied AI. Existing spatial audi

PhyLatent: Learning Dynamics-Relevant Representations for JEPA World Models

SafetyDGX agent

arXiv:2608.05720v1 Announce Type: new Abstract: We propose PhyLatent, a dynamics-relevant training objective for JointEmbedding Predictive Architecture (JEPA) world models. Our key observation is that

Physics-Based Molecular Fingerprints from Spectral Graph Theory Provide Efficient Geometry-Aware Measures of Chemical Similarity

ResearchDGX agent

arXiv:2608.05336v1 Announce Type: cross Abstract: Molecular representations are essential for the evaluation of molecular similarity and the development of structure-property relationships. Despite th

Pixel-TTS: Image based Text Rendering for Robust Text-to-Speech

ResearchDGX agent

arXiv:2606.14750v2 Announce Type: replace-cross Abstract: Recent advances in pixel-based text modeling show that representing text as images enables models to exploit visual cues for language understa

Plausible Patients, Impossible Populations: Auditing Epidemiological Fidelity in Large Language Model Mental Health Simulations

Model ReleasesDGX agent

arXiv:2604.17359v2 Announce Type: replace-cross Abstract: Language models asked to simulate psychiatric patients produce cases that survive inspection one at a time and populations that match no real

Poli-Bias: Understanding and Measuring Large Language Model Biases in International Political Conflicts

SafetyDGX agent

arXiv:2608.06123v1 Announce Type: new Abstract: Measuring political bias in large language models (LLMs) remains challenging as it can manifest through subtle differences in framing, argumentation, an

PolyAlign: Conditional Human-Distribution Alignment

SafetyDGX agent

arXiv:2606.13227v2 Announce Type: replace Abstract: Post-training methods such as supervised fine-tuning (SFT) and preference optimization typically align language models toward a single global assist

← Previous
1…5354555657…989
Next →