AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
Human
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
10 Jun 2026

Reasoning or Memorization? Direction-Aware Diversity Exploration in LLM Reinforcement Learning

SafetyDGX agent

arXiv:2606.10346v1 Announce Type: new Abstract: Reinforcement learning has become a key paradigm for eliciting reasoning abilities in large language models, where exploration is crucial for discoverin

Reasoning over Semantic IDs Enhances Generative Recommendation

SafetyDGX agent

arXiv:2603.23183v2 Announce Type: replace-cross Abstract: Recent advances in generative recommendation have leveraged pretrained LLMs by formulating sequential recommendation as autoregressive generat

Recalling Too Well: Sycophancy Evaluation and Mitigation in Memory-Augmented Models

Model ReleasesDGX agent

arXiv:2606.10949v1 Announce Type: new Abstract: Persistent memory systems promise to make LLMs more helpful by storing user beliefs over time. We show they also make models less correct by systematica

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Recoverable but Not Stationary:Local Linear Structures in Weights and Activations

Model ReleasesDGX agent

arXiv:2606.10929v1 Announce Type: cross Abstract: Task vectors, LoRA, activation steering, and random search around pretrained weights all suggest that learned behaviour can be controlled by linear di

Recovering the Zipfian Distribution in Unsupervised Term Discovery

SafetyDGX agent

arXiv:2606.10781v1 Announce Type: cross Abstract: Unsupervised term discovery involves segmenting unlabelled speech into word- or syllable-like units and clustering these into a lexicon of candidate t

RedAct: Redacting Agent Capability Traces for Procedural Skill Protection

Model ReleasesDGX agent

arXiv:2606.10813v1 Announce Type: cross Abstract: Users rely on execution traces to observe agent behavior, diagnose failures, and ensure accountability. These traces contain rich procedural detail, i

ReflectiChain: Epistemic Grounding in LLM-Driven World Models for Supply Chain Resilience

Model ReleasesDGX agent

arXiv:2606.10359v1 Announce Type: new Abstract: AI agents in supply chains face a fundamental epistemic gap: large language models (LLMs) interpret policies but lack physical grounding, while reinforc

Regimes: An Auditable, Held-Out-Gated Improvement Loop Demonstrated on LongMemEval with ActiveGraph

AgentsDGX agent

arXiv:2606.10241v1 Announce Type: new Abstract: Autonomous improvement loops are hard to trust because the improvement process is usually external scaffolding bolted onto the agent: failures go unlogg

Representation-Aware Advantage Estimation: Your Reward Model Provides More Than A Scalar Output

ResearchDGX agent

arXiv:2606.10528v1 Announce Type: cross Abstract: Current reinforcement learning from human feedback (RLHF) methods primarily rely on scalar rewards from a trained reward model (RM). While effective,

Representation Curriculum: Stagewise Training for Robust Ranking and Allocation

SafetyDGX agent

arXiv:2606.09891v1 Announce Type: cross Abstract: Ranking in digital marketplaces is a dynamic exposure-allocation mechanism: displayed items shape discovery trajectories and success events logged by

Resilient Navigation for Autonomous Farm Robots by Leveraging Jerk-Augmented Models with IMU-Only Disturbance Rejection

AgentsDGX agent

arXiv:2606.10971v1 Announce Type: new Abstract: Precise state estimation for navigation of autonomous agricultural robots is often compromised by sensor outages (GNSS/LiDAR/Visual) and high-frequency

Rethinking Embodied Navigation via Relational Inductive Bias

SafetyDGX agent

arXiv:2606.10348v1 Announce Type: new Abstract: Object navigation requires an agent to locate a target in an unknown environment through visual observations. Existing methods typically rely on open-vo

Rethinking the Flow-Based Gradual Domain Adaptation: A Semi-Dual Optimal Transport Perspective

ResearchDGX agent

arXiv:2602.01179v2 Announce Type: replace Abstract: Gradual domain adaptation (GDA) aims to mitigate domain shift by progressively adapting models from the source domain to the target domain via inter

Revisiting Metric Reliability for Fine-grained Evaluation of Machine Translation and Summarization in Indian Languages

Model ReleasesDGX agent

arXiv:2510.07061v2 Announce Type: replace Abstract: While automatic metrics drive progress in Machine Translation (MT) and Text Summarization (TS), existing metrics have been developed and validated a

Revisiting Positive Samples in Graph Contrastive Learning: From the Perspective of Message Passing

ResearchDGX agent

arXiv:2606.10284v1 Announce Type: new Abstract: Graph Contrastive Learning (GCL), which trains graph encoders by maximizing similarity between positive samples and minimizing it between negative ones,

Risk Comparisons in Linear Regression: Implicit Regularization Dominates Explicit Regularization

ResearchDGX agent

arXiv:2509.17251v2 Announce Type: replace-cross Abstract: Existing theory suggests that for linear regression problems categorized by capacity and source conditions, gradient descent (GD) is always mi

RKSC: Reasoning-Aware KV Cache Sharing and Confident Early Exit for Multi-Step LLM Inference

ResearchDGX agent

arXiv:2606.09937v1 Announce Type: cross Abstract: We introduce RKSC (Reasoning-Aware KV Cache Sharing), a training-free inference framework that eliminates two structural redundancies in multi-branch

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2510.14828v3 Announce Type: replace Abstract: Improving the reasoning capabilities of embodied agents is crucial for robots to complete complex human instructions in long-view manipulation tasks

RoboNaldo: Accurate, Stable and Powerful Humanoid Soccer Shooting via Motion-Guided Curriculum Reinforcement Learning

SafetyDGX agent

arXiv:2606.11092v1 Announce Type: cross Abstract: Elite humanoid soccer shooting requires whole-body stability, high-impulse whole-body interactions, and accuracy to targets. Motion tracking-driven re

RobotEQ: Transitioning from Passive Intelligence to Active Intelligence in Embodied AI

Model ReleasesDGX agent

arXiv:2605.06234v2 Announce Type: replace Abstract: Embodied AI is a prominent research topic in both academia and industry. Current research centers on completing tasks based on explicit user instruc

Robotic Nonprehensile Object Transportation with a Hanging Tray

ResearchDGX agent

arXiv:2606.10039v1 Announce Type: new Abstract: We consider the nonprehensile object transportation task known as the waiter's problem, in which a robot must move an object balanced on a tray from one

Robust Active Learning for Few-Shot Example Selection in Text-to-SQL

ResearchDGX agent

arXiv:2606.10125v1 Announce Type: cross Abstract: Few-shot example retrieval is the dominant paradigm for grounding large language models (LLMs) in domain-specific text-to-SQL systems. However, the qu

Robust Deep Reinforcement Learning Through Adversarial Attacks and Training : A Survey

AgentsDGX agent

arXiv:2403.00420v3 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) is a subfield of machine learning for training autonomous agents that take sequential actions across complex

Robust Regression of General ReLUs with Queries

SafetyDGX agent

arXiv:2606.11130v1 Announce Type: new Abstract: We study the task of agnostically learning general (as opposed to homogeneous) ReLUs under the Gaussian distribution with respect to the squared loss. I

Rod models in continuum and soft robot control: a review

ApplicationsDGX agent

arXiv:2407.05886v3 Announce Type: replace Abstract: Continuum and soft robots can transform automation tasks requiring compliant interaction in constrained or unstructured environments, including heal

Role-Agent: Bootstrapping LLM Agents via Dual-Role Evolution

SafetyDGX agent

arXiv:2606.10917v1 Announce Type: new Abstract: Although Large Language Model (LLM) agents have demonstrated strong performance on complex tasks, their learning is often limited by inefficient interac

ros2probe: Non-intrusive, Kernel-selective Observability for Robot Operating System 2 Middleware

ResearchDGX agent

arXiv:2606.10746v1 Announce Type: new Abstract: Robot Operating System 2 (ROS 2), the de facto standard middleware framework for robots, runs each robot as a graph of nodes communicating over the Data

Rotate2Think: Geometric Priming via Orthogonal Rotation to Improve Language Model Reasoning

Model ReleasesDGX agent

arXiv:2606.09873v1 Announce Type: cross Abstract: Reasoning models achieve strong performance on challenging tasks by generating explicit intermediate reasoning traces before producing a final answer.

Routing-Aware Expert Calibration for Machine Unlearning in Mixture-of-Experts Language Models

ResearchDGX agent

arXiv:2606.10338v1 Announce Type: cross Abstract: Machine unlearning is increasingly important for large language models, yet unlearning in Mixture-of-Experts (MoE) architectures remains underexplored

SAFE: An LLM-as-Verifier Framework for Evidence-Grounded Multi-Hop Reasoning

Model ReleasesDGX agent

arXiv:2604.01993v2 Announce Type: replace-cross Abstract: Multi-hop QA benchmarks often reward Large Language Models (LLMs) for spurious correctness, where models reach correct answers through invalid

Sample Where You Struggle: Sharpening Base Model Reasoning via Entropy-Guided Power Sampling

Model ReleasesDGX agent

arXiv:2606.09926v1 Announce Type: cross Abstract: Sampling from the sequence-level power distribution p^alpha elicits RL-level reasoning from base language models without any parameter updates, but th

Sampling Triangulations and Calabi-Yau Threefolds with Autoregressive GNNs

HardwareDGX agent

arXiv:2605.27770v2 Announce Type: cross Abstract: We introduce `dualGNN', an autoregressive message-passing GNN for sampling fine, regular triangulations (FRTs) of convex polytopes. dualGNN operates o

SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation

Model ReleasesDGX agent

arXiv:2606.10305v1 Announce Type: new Abstract: Fine-tuning vision-language-action (VLA) policies for long-horizon manipulation still relies heavily on behavior cloning, which requires costly high-qua

SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning

Model ReleasesDGX agent

arXiv:2606.10804v1 Announce Type: new Abstract: Controlled character animation requires transferring motion from a driving sequence to a reference character. Prior works heavily rely on intermediate r

Scaling Self-Supervised Speech Models Uncovers Deep Linguistic Relationships: Evidence from the Pacific Cluster

ResearchDGX agent

arXiv:2603.07238v2 Announce Type: replace Abstract: Similarities between language representations derived from Self-Supervised Speech Models (S3Ms) have been observed to primarily reflect geographic p

Schmidt Decomposition-Based Methods for Efficient Quantum Image Encoding

ResearchDGX agent

arXiv:2606.10874v1 Announce Type: new Abstract: In quantum image processing, a fundamental step is encoding classical image data into quantum states. This can be achieved using methods such as Flexibl

SCOPE: Sequential Causal Optimization of Process Interventions

Model ReleasesDGX agent

arXiv:2512.17629v4 Announce Type: replace-cross Abstract: Prescriptive Process Monitoring (PresPM) recommends interventions during running business processes to optimize key performance indicators (KP

SD-GRPO: Verifiable Segment Decomposition for Long-Form Vision-Language Generation

SafetyDGX agent

arXiv:2606.09871v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) and its variants, originally developed for Large Language Models (LLMs), have recently been applied to Multi

Seal-Robust KCR: A Robust Kuzushiji Character Recognition Framework under Seal Interference

ResearchDGX agent

arXiv:2602.19086v2 Announce Type: replace Abstract: Kuzushiji was one of the most widely used cursive writing systems in pre-modern Japan. Due to its highly cursive forms and extensive glyph variation

Secure Aggregation with Top-K Sparsification in Decentralized Federated Learning

ResearchDGX agent

arXiv:2606.10780v1 Announce Type: cross Abstract: Secure aggregation is a vital component for mitigating gradient leakage in federated learning, but its communication cost conventionally scales with t

Segment and Select: Vision-Language Segmentation in 3D Scenarios

ResearchDGX agent

arXiv:2606.10594v1 Announce Type: new Abstract: 3D vision-language segmentation aims to segment target objects in 3D scenarios according to the linguistic instructions and visual observations. Prior a

Segmentation-Driven Monocular Shape from Polarization based on Physical Model

ApplicationsDGX agent

arXiv:2601.04776v2 Announce Type: replace Abstract: Monocular shape-from-polarization (SfP) leverages the intrinsic relationship between light polarization properties and surface geometry to recover s

Selection, Not Salience: The Shape and Limits of Personalization in Social Highlighting

SafetyDGX agent

arXiv:2606.10398v1 Announce Type: cross Abstract: Does personalizing what a reader sees pay off, and where does it stop? Using a social web highlighter and a co-readership identity control (the same d

Selective Disk Bispectrum: A Complete and Rotation Invariant Image Descriptor

SafetyDGX agent

arXiv:2511.19706v2 Announce Type: replace-cross Abstract: Rotation invariance is a fundamental requirement across many computer vision tasks. Historically, this inductive bias has been encoded through

Self-Distillation Policy Optimization via Visual Feedback: Bridging Code and Visual Artifacts

SafetyDGX agent

arXiv:2606.10334v1 Announce Type: new Abstract: Code-generating large language models (LLMs) increasingly produce visual artifacts such as charts, web pages, and slides by writing programs that are ex

Self-EmoQ: Plutchik-Guided Value-based Planning to Drive Streaming Emotional TTS

SafetyDGX agent

arXiv:2606.09837v1 Announce Type: cross Abstract: Emotional interaction is increasingly crucial for conversational AI, yet current systems lack a self-emotion determination mechanism to drive the stre

Self-Supervised Relevance Modelling in Autonomous Driving via Counterfactual Analysis

SafetyDGX agent

arXiv:2606.10688v1 Announce Type: new Abstract: Autonomous driving relies on computationally intensive perception pipelines to continuously detect and track objects in the surrounding environment. Whi

SHAPE: Coalition-Aware Expert Pruning for Sparse Mixture-of-Experts LLMs

Model ReleasesDGX agent

arXiv:2606.09886v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) large language models achieve strong quality with low per-token compute, yet their deployment is often limited by the

SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration

Model ReleasesDGX agent

arXiv:2606.10228v1 Announce Type: cross Abstract: Safe exploration is a prerequisite for deploying reinforcement learning (RL) agents in safety-critical domains. In this paper, we approach safe explor

Sigma-Branch: Hierarchical Single-Path Network Reconstruction for Dynamic Inference with Reduced Active Parameters

Model ReleasesDGX agent

arXiv:2606.09924v1 Announce Type: cross Abstract: Deploying deep neural networks on memory-constrained edge accelerators is bottlenecked by per-inference off-chip weight transfer rather than computati

Sim2Schedule: A Simulator-Guided LLM Framework for Autonomous Open-Pit Mine Scheduling

Model ReleasesDGX agent

arXiv:2606.10286v1 Announce Type: new Abstract: Open-pit mine scheduling is a critical process for maximizing economic return under complex geotechnical and operational constraints. While Mixed-Intege

SinkRec: Mitigating Semantic State Sink in Long Sequence Recommendation with Memory-Conditioned Gated Delta Networks

SafetyDGX agent

arXiv:2606.09888v1 Announce Type: new Abstract: Linear attention provides an efficient backbone for long-sequence recommendation by avoiding the quadratic cost of standard Transformers, but its compre

Sketch-to-Layout: A Human-Centric Computational Agent for Constraint-Aware Synthesis of Modular Photobioreactors

SafetyDGX agent

arXiv:2606.09849v1 Announce Type: cross Abstract: Building-integrated photobioreactors (PBRs) offer a pathway for carbon-neutral architecture, yet deployment is hindered by configuration complexity an

SkillResolve-Bench: Measuring and Resolving Same-Capability Ambiguity in Agent Skill Retrieval

Model ReleasesDGX agent

arXiv:2606.10388v1 Announce Type: cross Abstract: Agent skill libraries are becoming routable software assets: a retrieved skill can contribute instructions, scripts, resource bindings, and execution

Sleep EEG Signal Criticality as a Non-Invasive Predictor of Cognitive Decline in Dementia

ResearchDGX agent

arXiv:2606.10889v1 Announce Type: cross Abstract: Early detection of neurodegeneration remains a critical clinical challenge. This study investigates whether sleep EEG signal criticality, quantified v

Small Data, Big Noise: Adversarial Training for Robust Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.10610v1 Announce Type: new Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become essential for adapting foundation models to downstream NLP tasks. However, current PEFT methods often

SocraticPO: Policy Optimization via Interactive Guidance

SafetyDGX agent

arXiv:2606.09887v1 Announce Type: cross Abstract: Reinforcement learning (RL) for large language models usually supervises reasoning with scalar outcome rewards, such as binary correctness. Such rewar

SoK: Colluding Adversaries in Machine Learning Pipelines

SafetyDGX agent

arXiv:2606.10091v1 Announce Type: cross Abstract: Machine learning (ML) models are susceptible to various security, privacy, and fairness risks. Adversaries with different characteristics (i.e., objec

Soul Computing: A Theoretical Framework and Technical Architecture for Intelligent Agents with Independent Consciousness

ResearchDGX agent

arXiv:2606.10413v1 Announce Type: new Abstract: Breakthroughs in large language models and multimodal generation technologies have propelled the digital reconstruction of human mental traits, emotiona

SPACE: Source-free Proxy Anchor Concept Erasure for MLLMs

Model ReleasesDGX agent

arXiv:2606.09868v1 Announce Type: cross Abstract: As Multimodal Large Language Models (MLLMs) face growing privacy risks and regulatory constraints, machine unlearning (MU) has emerged as a crucial so

← Previous
1…433434435436437…1049
Next →