AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
Human
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
64,475 results
Research

SHAP-Guided Kernel Actor-Critic for Explainable Reinforcement Learning

DGX agent

arXiv:2512.05291v3 Announce Type: replace Abstract: Actor-critic (AC) methods are a cornerstone of reinforcement learning (RL) but offer limited interpretability. Current explainable RL methods seldom

researcharxiv-cs-lg
8 Jun 2026
Safety
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Shield-Loco: Shielding Locomotion Policies with Predictive Safety Filtering

DGX agent

arXiv:2606.07193v1 Announce Type: new Abstract: Reinforcement learning (RL) policies enable dynamic legged locomotion but lack mechanisms to avoid violations of safety constraints that are absent duri

safetyarxiv-cs-ro
8 Jun 2026
Agents

Should You Use Your Large Language Model to Explore or Exploit?

DGX agent

arXiv:2502.00225v4 Announce Type: replace-cross Abstract: We evaluate the ability of the current generation of large language models (LLMs) to help a decision-making agent facing an exploration-exploi

agentsarxiv-cs-ai
8 Jun 2026
Model Releases

SigmaScale: LLM Compression with SVD-based Low-Rank Decomposition and Learned Scaling Matrices

DGX agent

arXiv:2606.07098v1 Announce Type: new Abstract: We present SigmaScale, a method for learning auxiliary scaling matrices S to aid truncated Singular Value Decomposition (SVD) based Large Language Model

model-releasesarxiv-cs-cl
8 Jun 2026
Agents

Signal-Driven Observation for Long-Horizon Web Agents

DGX agent

arXiv:2606.06708v1 Announce Type: new Abstract: Web agents operating over long horizons ingest raw DOM and accessibility trees -- routinely tens of thousands of tokens -- at every action step, causing

agentsarxiv-cs-cl
8 Jun 2026
Safety

Simulation-Driven Imitation Learning for Biosignals-Free Shared-Autonomy Prosthetic Grasping

DGX agent

arXiv:2606.07389v1 Announce Type: new Abstract: Biosignals-free shared-autonomy control of upper-limb prosthetic hands aims to enable natural and low-effort manipulation without relying on EMG or othe

safetyarxiv-cs-ro
8 Jun 2026
Model Releases

Skill-3D: Evolving Scene-Aware Skills for Agentic 3D Spatial Reasoning

DGX agent

arXiv:2606.07436v1 Announce Type: new Abstract: This paper explores agentic 3D spatial understanding, i.e., MLLM agents performing 3D reasoning through tool use. Existing methods often misuse tools an

model-releasesarxiv-cs-cv
8 Jun 2026
Research

Skip a Layer or Loop It? Learning Program-of-Layers in LLMs

DGX agent

arXiv:2606.06574v1 Announce Type: new Abstract: Large language models (LLMs) perform inference by following a fixed depth and order, non-recurrent execution of all layers. We reveal the wide existence

researcharxiv-cs-lg
8 Jun 2026
Research

SleepExplain: Explainable Non-Rapid Eye Movement and Rapid Eye Movement Sleep Stage Classification from EEG Signal

DGX agent

arXiv:2606.07351v1 Announce Type: cross Abstract: Classification of sleep stages is one of the most important diagnostic approaches for a variety of sleep-related disorders. Electroencephalography (EE

researcharxiv-cs-ai
8 Jun 2026
Safety

SlimSearcher: Training Efficiency-Aware Web Agents via Adaptive Reward Gating

DGX agent

arXiv:2606.07074v1 Announce Type: cross Abstract: Deep research agents have demonstrated remarkable capabilities in complex information-seeking tasks, yet this power comes at a steep computational cos

safetyarxiv-cs-ai
8 Jun 2026
Agents

Small Language Model Agents Enable Efficient and High-Quality Knowledge Mining

DGX agent

arXiv:2510.01427v3 Announce Type: replace Abstract: At the core of Deep Research is knowledge mining, the task of extracting structured information from massive unstructured text in response to user i

agentsarxiv-cs-ai
8 Jun 2026
Safety

Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills

DGX agent

arXiv:2606.07412v1 Announce Type: cross Abstract: LLM-driven software engineering agents have become a central testbed for real-world language-model capability, yet their training remains limited by t

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

Sparse Subspace-to-Expert Sharing for Task-Agnostic Continual Learning

DGX agent

arXiv:2606.07500v1 Announce Type: cross Abstract: Continual learning in Large Language Models (LLMs) is hindered by the plasticity-stability dilemma, where acquiring new capabilities often leads to ca

model-releasesarxiv-cs-ai
8 Jun 2026
Applications

Sparsely gated tiny linear experts

DGX agent

arXiv:2606.07414v1 Announce Type: new Abstract: Sparsity allows scaling model parameters without proportionally increasing computational cost. While mixture of experts (MoE) models are made increasing

applicationsarxiv-cs-lg
8 Jun 2026
Model Releases

Spatial-Temporal Decoupled Adapter for Micro-gesture Online Recognition

DGX agent

arXiv:2606.07355v1 Announce Type: new Abstract: Micro-gesture online recognition aims to temporally localize and classify subtle gestures in untrimmed videos. Owing to their extremely short duration,

model-releasesarxiv-cs-cv
8 Jun 2026
Applications

Spatiotemporal Imputation with Graph-Informed Flow Matching

DGX agent

arXiv:2606.06682v1 Announce Type: new Abstract: Missing data is a common challenge in spatiotemporal systems, arising in applications such as air quality monitoring and urban traffic management. Tradi

applicationsarxiv-cs-lg
8 Jun 2026
Applications

SpectCount: Spectrotemporal Counting via Synthetic Signals Improves Large Audio Language Models

DGX agent

arXiv:2606.06907v1 Announce Type: cross Abstract: Large audio language models (LALMs) extend large language models with an audio encoder and large-scale audio data. However, the scarcity of high-quali

applicationsarxiv-cs-ai
8 Jun 2026
Model Releases

Spline Policy: A Structured Representation for Robot Policies

DGX agent

arXiv:2606.07386v1 Announce Type: new Abstract: Modern imitation-learning policies for robot manipulation often represent actions as fixed-resolution action chunks, which are simple and effective but

model-releasesarxiv-cs-ro
8 Jun 2026
Tutorials

SS-TPT: Stability and Suitability-Guided Test-Time Prompt Tuning for Adversarially Robust Vision-Language Models

DGX agent

arXiv:2606.06943v1 Announce Type: cross Abstract: Vision-language models (VLMs) such as CLIP achieve strong zero-shot recognition but remain highly fragile under adversarial perturbations. Recent test

tutorialsarxiv-cs-ai
8 Jun 2026
Research

Stability beyond Bounded Differences: Sharp Generalization Bounds under Finite L_p Moments

DGX agent

arXiv:2606.06855v1 Announce Type: cross Abstract: While algorithmic stability is a central tool for understanding generalization of learning algorithms, existing high-probability guarantees typically

researcharxiv-cs-lg
8 Jun 2026
Safety

Stable Reasoning, Unstable Responses: Mitigating LLM Deception via Stability Asymmetry

DGX agent

arXiv:2603.26846v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) expand in capability and application scope, their trustworthiness becomes critical. A vital risk is intrinsic

safetyarxiv-cs-ai
8 Jun 2026
Local Ai

StainFlow: Entity-Stain Tracking and Evidence Linking for Process Rewards in GUI Agents

DGX agent

arXiv:2606.07027v1 Announce Type: new Abstract: Reinforcement Learning (RL) has become a promising approach for improving GUI Agents in long-horizon, stochastic digital environments, but trajectory-le

local-aiarxiv-cs-ai
8 Jun 2026
Applications

Standard vs. Modular Sampling: Best Practices for Reliable LLM Unlearning

DGX agent

arXiv:2509.05316v2 Announce Type: replace-cross Abstract: A conventional LLM Unlearning setting consists of two subsets -'forget' and 'retain', with the objectives of removing the undesired knowledge

applicationsarxiv-cs-ai
8 Jun 2026
Safety

Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models

DGX agent

arXiv:2602.02600v3 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) have recently emerged as a competitive alternative to autoregressive (AR) models, offering parallel decoding,

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

STREAM: Stochastic Riemannian Flow Matching with Anisotropic Decoder for Digital Histopathology Image Generation

DGX agent

arXiv:2606.07036v1 Announce Type: cross Abstract: Synthetic histopathology image generation addresses critical challenges in computational pathology, including patient privacy and the growing need for

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Stream3D-VLM: Online 3D Spatial Understanding with Incremental Geometry Priors

DGX agent

arXiv:2606.06891v1 Announce Type: new Abstract: Despite advances in 3D scene understanding, existing 3D Large Multimodal Models operate in offline settings, requiring complete scene observations or pr

model-releasesarxiv-cs-cv
8 Jun 2026
Local Ai

Streaming Video Generation with Streaming Force Control

DGX agent

arXiv:2606.07508v1 Announce Type: new Abstract: We introduce StreamForce, a streaming video generation framework that enables physically grounded control through continuous force inputs. Unlike prior

local-aiarxiv-cs-cv
8 Jun 2026
Research

STRIPS-WM: Learning Grounded Propositional STRIPS-style World Models from Images

DGX agent

arXiv:2606.06832v1 Announce Type: new Abstract: Robots performing long-horizon visual manipulation observe high-dimensional images, but successful plans depend on action-relevant facts: what can be do

researcharxiv-cs-ro
8 Jun 2026
Tutorials

Structure-Preserving Correction Learning for Sparse Bayesian Inference in Brain Source Imaging

DGX agent

arXiv:2606.07196v1 Announce Type: new Abstract: Classical sparse Type-II Bayesian methods for M/EEG brain imaging support joint estimation of source and noise hyperparameters, but rely on fixed iterat

tutorialsarxiv-cs-lg
8 Jun 2026
Model Releases

Style or Content? Evaluating Style Classifiers with Controlled Content Overlap

DGX agent

arXiv:2606.07103v1 Announce Type: new Abstract: Style classifiers can use content cues that correlate with style labels in naturally collected data, yet we lack a systematic way to measure this relian

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

Superintelligent Retrieval Agent: The Next Frontier of Agentic Retrieval

DGX agent

arXiv:2605.06647v2 Announce Type: replace-cross Abstract: Retrieval-augmented agents are increasingly the interface to large knowledge bases, yet most treat retrieval as a black box: they issue explor

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Supervision versus Demonstration-Based In-Context Learning for Multiword Expression Classification

DGX agent

arXiv:2606.07479v1 Announce Type: cross Abstract: Turkish idiomatic light verb constructions (LVCs) are challenging for multiword expression processing because they often share the same surface form a

model-releasesarxiv-cs-ai
8 Jun 2026
Safety

SV-Detect: AI-generated Text Detection with Steering Vectors

DGX agent

arXiv:2606.07313v1 Announce Type: cross Abstract: Detecting machine-generated text is especially difficult under distribution shift, such as transfer across domains, source models, and editing attacks

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

SVHighlights: Towards Extremely Long Sport Video Highlight Detection

DGX agent

arXiv:2606.06926v1 Announce Type: new Abstract: While highlight detection for long-form videos is of great practical importance, most existing methods remain limited to short-form content, largely due

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

SW-A^2-Bench: Benchmarking Autonomous Software Agent Generation for Agentic Web

DGX agent

arXiv:2604.04226v2 Announce Type: replace-cross Abstract: The Agentic Web is emerging as a paradigm in which autonomous software agents interact with online resources and with each other to accomplish

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

SWE-Explore: Benchmarking How Coding Agents Explore Repositories

DGX agent

arXiv:2606.07297v1 Announce Type: cross Abstract: Repository-level coding benchmarks such as SWE-bench have driven a rapid surge in the capabilities of coding agents. Yet they usually treat coding tas

model-releasesarxiv-cs-cl
8 Jun 2026
Research

SWE-IF: Aligning Code Evaluation with Human Preference

DGX agent

arXiv:2510.07315v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have catalyzed vibe coding, where users leverage LLMs to generate and iteratively refine code through natural lan

researcharxiv-cs-ai
8 Jun 2026
Safety

Sycophantic Praise: Evaluating Excessive Praise in Language Models

DGX agent

arXiv:2606.07441v1 Announce Type: new Abstract: Sycophancy in language models is typically studied as excessive agreement or validation, while explicit praise and flattery have received comparatively

safetyarxiv-cs-cl
8 Jun 2026
Research

Synthetic Benchmarks Overstate Forward-Forward Scaling: Real-Data Limits of Layer-Local Training

DGX agent

arXiv:2606.06539v1 Announce Type: cross Abstract: Forward-Forward (FF) learning [Hinton, 2022] replaces backpropagation with strictly layer-local goodness updates. Recent FF-CNN work has narrowed the

researcharxiv-cs-ai
8 Jun 2026
Applications

Synthics: Synthetic Physics-like Datasets for Machine Learning

DGX agent

arXiv:2606.06724v1 Announce Type: new Abstract: Representative data is fundamental in machine learning, as limited data hinders generalisation. Collecting sufficient real-world samples is often infeas

applicationsarxiv-cs-lg
8 Jun 2026
Safety

T-GMP: Terrain-conditioned Generative Motion Priors for Versatile and Natural Humanoid Locomotion

DGX agent

arXiv:2606.06944v1 Announce Type: new Abstract: Achieving both anthropomorphic naturalness and robust terrain traversal remains a fundamental challenge in humanoid locomotion. Existing Reinforcement L

safetyarxiv-cs-ro
8 Jun 2026
Research

T2LM: Long-Term 3D Human Motion Generation from Multiple Sentences

DGX agent

arXiv:2406.00636v2 Announce Type: replace Abstract: In this paper, we address the challenging problem of long-term 3D human motion generation. Specifically, we aim to generate a long sequence of smoot

researcharxiv-cs-cv
8 Jun 2026
Research

TA-RAG: Tone-Aware Retrieval-Augmented Generation for Peer-Support Health Communication

DGX agent

arXiv:2606.06794v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) successfully grounds large language model (LLM) outputs in trusted documents, but factual grounding alone is insuff

researcharxiv-cs-cl
8 Jun 2026
Research

TabSwift: An Efficient Tabular Foundation Model with Row-Wise Attention

DGX agent

arXiv:2606.07345v1 Announce Type: new Abstract: Tabular foundation models, exemplified by TabPFN, perform prediction via in-context learning, inferring test labels directly from labeled training examp

researcharxiv-cs-lg
8 Jun 2026
Model Releases

TALAN: Task-Aligned Latent Adaptation Networks for Targeted Post-Training of Large Language Models

DGX agent

arXiv:2606.06902v1 Announce Type: new Abstract: Targeted post-training aims to improve reasoning, math, and code without degrading strengths. Low-rank adapters are efficient but task-global; activatio

model-releasesarxiv-cs-lg
8 Jun 2026
Applications

TargetSEC: Plug-and-Play In-the-Wild Speech Emotion Conversion via Arousal-Conditioned Latent Style Diffusion

DGX agent

arXiv:2606.07293v1 Announce Type: cross Abstract: Speech Emotion Conversion (SEC) aims to transform the emotion of a source utterance into a target emotion while preserving content and speaker identit

applicationsarxiv-cs-lg
8 Jun 2026
Local Ai

Task Editing for Generalizable 3D Visuomotor Policy Learning

DGX agent

arXiv:2606.07012v1 Announce Type: new Abstract: 3D visuomotor policies offer a promising direction for complex robotic manipulation, as depth maps and point clouds provide rich geometric information f

local-aiarxiv-cs-ro
8 Jun 2026
Safety

Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling

DGX agent

arXiv:2507.06419v3 Announce Type: replace Abstract: Reward modeling (RM), which captures human preferences to align large language models (LLMs), is increasingly employed in tasks such as model finetu

safetyarxiv-cs-cl
8 Jun 2026
← Previous
1…615616617618619…1344
Next →