AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
30 Jul 2026

SeasonStereo: Robust Dense Stereo Matching for Multi-Date Satellite Imagery via Generative AI

ResearchDGX agent

arXiv:2607.27139v1 Announce Type: new Abstract: Accurate 3D reconstruction from satellite imagery typically relies on near-simultaneous stereo pairs, limiting its applicability to diachronic settings

SecRespond: Benchmarking AI Agents for Real-World Post-Compromise Incident Response

Model ReleasesDGX agent

arXiv:2607.26791v1 Announce Type: cross Abstract: Large Language Model (LLM) agents are increasingly adopted in real-world security operations with access to host artifacts and command-line interfaces

See2Think: Do Multimodal Models Really Use Intermediate Visual States?

ApplicationsDGX agent

arXiv:2607.26769v1 Announce Type: new Abstract: Multimodal large language models increasingly use sketches, annotations, tools, and intermediate images during reasoning, but it remains unclear whether

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Seeing or Knowing? Visual Context Sensitivity in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.26326v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) achieve strong performance by integrating visual inputs with the rich priors of pretrained language models. How

Self-Adaptive Learning and Model Predictive Control for Tracking Unknown Dynamics with No Regret

SafetyDGX agent

arXiv:2607.26370v1 Announce Type: cross Abstract: We propose a self-adaptive online learning for control method for tracking unknown target dynamics. The target dynamics can exhibit switching behavior

Self-Configurable Mesh-Networks for Scalable Distributed Submodular Bandit Optimization

AgentsDGX agent

arXiv:2602.19366v2 Announce Type: replace-cross Abstract: We study how to scale distributed bandit submodular coordination under realistic communication constraints in bandwidth, data rate, and connec

Semantic-Aware Temporal Adaptation for UAV Anti-UAV Tracking

Model ReleasesDGX agent

arXiv:2607.26511v1 Announce Type: new Abstract: UAV Anti-UAV tracking is an emerging low-altitude security task for localizing an adversarial UAV using the onboard camera of a moving observer UAV. It

Semi-Decentralized Multi-Spacecraft Collision Avoidance under Communication Constraints

AgentsDGX agent

arXiv:2607.26570v1 Announce Type: new Abstract: Current spacecraft collision-avoidance operations rely on intermittent ground-station contacts, requiring operators to plan with delayed and asynchronou

SENSE: Efficient EEG-to-Text via Privacy-Preserving Semantic Retrieval

Local AiDGX agent

arXiv:2603.17109v2 Announce Type: replace Abstract: Decoding brain activity into natural language is a major challenge in AI with important applications in assistive communication, neurotechnology, an

Sensor-Placement-Agnostic Sonomyography: Toward Continuous High-Dimensional Control by Users with Tetraplegia

ResearchDGX agent

arXiv:2607.26401v1 Announce Type: cross Abstract: Sonomyography (SMG) enables continuous device control via ultrasound-measured muscle deformation signals, but existing SMG interfaces generally requir

Sequence-SOD: Bio-inspired Sequence-aware Spiking ObjectDetection for Event Cameras

Local AiDGX agent

arXiv:2607.26703v1 Announce Type: new Abstract: Event cameras follow a retina-inspired sensing principle, reporting local intensity changes asynchronously with hightemporal resolution and a wide dynam

SERPO: Self-Evolving Rubric Policy Optimization for Open-Ended Test-Time Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.26873v1 Announce Type: new Abstract: Test-time reinforcement learning (TTRL) enables language models to self-evolve at inference time without labeled feedback. Existing methods rely on answ

Setoka: A Benchmark for Hierarchical User Understanding in Personalized Agents over Heterogeneous Data

Model ReleasesDGX agent

arXiv:2607.27056v1 Announce Type: cross Abstract: Personalized agents are increasingly applied to assist users across a wide range of tasks. Effective personalized assistance requires not only retriev

SG-CoT: An Ambiguity-Aware Robotic Planning Framework using Scene Graph Representations

AgentsDGX agent

arXiv:2603.18271v3 Announce Type: replace Abstract: Ambiguity poses a major challenge to large language models (LLMs) used as robotic planners. In this letter, we present Scene Graph-Chain-of-Thought

Shape-Based Inductive Bias for Glioma Grading from Tumor Contours

SafetyDGX agent

arXiv:2607.26090v1 Announce Type: cross Abstract: Glioma grading from tumor contours is often treated as a pixel problem even when the signal of interest is shape. We align closed contours with a func

Shared SFT Lessons Across Alignment, Model Organisms, and Toy Models

Model ReleasesDGX agent

arXiv:2607.26173v1 Announce Type: new Abstract: Alignment training, model organisms, and toy models are usually treated as separate research areas. But projects in all three frequently use supervised

Shot-based quantum encoding: a data-loading paradigm for quantum neural networks

ResearchDGX agent

arXiv:2604.06135v2 Announce Type: replace-cross Abstract: Efficient data loading remains a bottleneck for near-term quantum machine learning. Existing schemes (angle, amplitude, and basis encoding) ei

Sim2Win: A Team-Agnostic, Event-Based Pre-Match Outcome Prediction and Tactical Profiling System for Football

ResearchDGX agent

arXiv:2607.26061v1 Announce Type: new Abstract: Pre-match tactical decision-making in professional football relies heavily on subjective expert analysis and identity-based scouting systems that cannot

Simplex Demixing: Disentangling Multiple Light-Flavor Jets at Colliders

ApplicationsDGX agent

arXiv:2607.24921v1 Announce Type: cross Abstract: Providing a practical and hadron-level definition of multiple jet flavors has been a long-standing challenge in collider physics. Previous work has in

Simultaneous Coverage and Efficiency Guarantee in Online Conformal Prediction

Model ReleasesDGX agent

arXiv:2607.26577v1 Announce Type: new Abstract: Adaptive conformal inference (ACI) of Gibbs and Cand{es and its variants are the standard approach to online conformal prediction under distribution shi

Single-Beat Cuffless Blood Pressure Estimation Using Ear-PPG and ECG with a Lightweight Hybrid Learning Framework

ApplicationsDGX agent

arXiv:2607.27076v1 Announce Type: new Abstract: Continuous cuffless blood pressure (BP) monitoring remains challenging due to motion artifacts, physiological variability, and the limited robustness of

SkillCAT: Contrastive, Assessment-Augmented and Topology-AwareSkill Self-Evolution for LLM Agents

AgentsDGX agent

arXiv:2606.13317v2 Announce Type: replace Abstract: Skill self-evolution methods for LLM agents aim to turn execution trajectories into reusable skill documents. However, current pipelines typically d

Skillful forecasting of offshore winds from satellite scatterometer constellations

ResearchDGX agent

arXiv:2607.27152v1 Announce Type: new Abstract: Accurate intraday forecasts of offshore wind are becoming increasingly important for power system operation and the integration of growing shares of off

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution

SafetyDGX agent

arXiv:2607.26784v1 Announce Type: new Abstract: Large language model agents often encounter related yet distinct tasks that share reusable solution patterns. Yet standard agentic reinforcement learnin

Sky sphere representation in language models

ResearchDGX agent

arXiv:2607.27092v1 Announce Type: new Abstract: We analyze whether language models of size ~100B have a representation of the night sky map that is decodable from their residual stream. We find that m

SMSP: A Plug-and-Play Strategy of Multi-Scale Perception for MLLMs to Perceive Visual Illusions

SafetyDGX agent

arXiv:2603.23118v2 Announce Type: replace Abstract: Recent works have shown that multimodal large language models (MLLMs) are highly vulnerable to hidden-pattern visual illusions, where the hidden con

SpatialQ: Understanding 3D Gaussian Splatting Scene Quality via Visual-based MLLM

Model ReleasesDGX agent

arXiv:2607.26595v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has emerged as an effective representation for novel view synthesis and 3D scene reconstruction, creating an increasing dem

SpecFirst: Behavioral Specification Elicitation as a First-Class Step in Agent-Based Program Synthesis from Scratch

AgentsDGX agent

arXiv:2607.27167v1 Announce Type: cross Abstract: LLM-based agents excel at software engineering tasks where an existing codebase provides context, but constructing a program from scratch remains fund

Speech2Grasp: Data-Efficient Transfer of Text-Conditioned Grasp Detection to Speech in Humanoid Robots

ApplicationsDGX agent

arXiv:2607.26567v1 Announce Type: cross Abstract: Humanoid robots increasingly require multi-modal understanding for natural interaction with humans. Despite the prominence of vision-language models,

Spline-Based Boundary Representations for Sparse View Reconstruction and Simulation Using Isogeometric Analysis

ResearchDGX agent

arXiv:2607.26234v1 Announce Type: new Abstract: Image-based reconstruction aims to recover three-dimensional geometry from images. Recent advances have enabled the recovery of visually detailed models

SPROUT: A Scalable Diffusion Foundation Model for Agricultural Vision

ResearchDGX agent

arXiv:2603.27519v2 Announce Type: replace Abstract: Image-based plant phenotyping depends on dense structural understanding of crops, yet pixel-level annotation remains expensive across species, organ

Stable and Budget-Feasible Coalition Formation for Clustered Federated Learning: A Hedonic Potential-Game Approach

SafetyDGX agent

arXiv:2607.26788v1 Announce Type: cross Abstract: Clustered federated learning benefits from organizing heterogeneous participants into coalitions that train coalition-specific models, but such cluste

StatePlay: State-Aware Game World Models for Mechanics-Consistent Generation

TutorialsDGX agent

arXiv:2607.26754v1 Announce Type: new Abstract: Recent game world models can generate visually realistic and interactive environments conditioned on player actions. However, games are not defined by p

Statistical laws and linguistics differ in naturalistic video and fictional conversations

ResearchDGX agent

arXiv:2512.18072v3 Announce Type: replace Abstract: Conversation is a cornerstone of social connection and is linked to well-being outcomes. Conversations vary widely in type with some portion generat

Steering Instruction Hierarchies at Inference Time

SafetyDGX agent

arXiv:2607.26228v1 Announce Type: new Abstract: Instruction hierarchies are a core safety assumption of language model deployment: higher priority inputs, such as system prompts, should override confl

Step-Attention Refinement of DINOv3 Features for Efficient Anterior Eye Segmentation

ResearchDGX agent

arXiv:2607.27087v1 Announce Type: new Abstract: Anterior eye segment (AES) segmentation is a key component of both ocular biometrics and emerging clinical image analysis applications. However, heterog

Structurally Separated Uncertainty in Supervised Latent Variable Models

Model ReleasesDGX agent

arXiv:2602.11219v2 Announce Type: replace Abstract: Predictive uncertainty is commonly decomposed into epistemic and aleatoric components, but standard decompositions often produce strongly correlated

StructureGS: Structure-aware Gaussian Splatting for Articulated Object Reconstruction

ResearchDGX agent

arXiv:2607.26889v1 Announce Type: cross Abstract: Reconstructing articulated objects with multiple movable parts is essential for understanding object structure and enabling physical interaction. Howe

Surrogate assisted diversity estimation in neural ensemble search

TutorialsDGX agent

arXiv:2607.26940v1 Announce Type: new Abstract: Ensembles are a standard way to improve the performance and robustness of deep neural networks, but their effectiveness crucially depends on both the qu

SymmGrid: Super-Scaling On-Robot Learning with Parallelized Symmetries and Egocentric-Exocentric Visual Perception

SafetyDGX agent

arXiv:2607.26985v1 Announce Type: cross Abstract: Deep reinforcement policy learning directly in physical robots (on-robot learning) remains bottlenecked by slow wall-clock training times. We present

Symphony of Bias: Exploring Gender Associations with Musical Instruments in Multimodal LLMs

Model ReleasesDGX agent

arXiv:2607.26355v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly embedded in everyday life and widely used for information seeking, raising concerns about their potential

Task and Skill Planning: Hierarchical Robot Planning with Black-Box Skills

ApplicationsDGX agent

arXiv:2504.17901v3 Announce Type: replace Abstract: Task and motion planning (TAMP) is a well-established approach for solving long-horizon robot planning problems. Although TAMP methods have historic

Temporally Centered SIGReg Improves Multi-Task LeWorldModel Learning: From Analysis to Method

Model ReleasesDGX agent

arXiv:2607.26924v1 Announce Type: new Abstract: Recent work on LeWorldModel (LeWM) has shown that the Sketched Isotropic Gaussian Regularizer (SIGReg) enables stable end-to-end world-model learning fr

The Advantage of Fine-Grained Training

SafetyDGX agent

arXiv:2509.05130v2 Announce Type: replace Abstract: In classification problems, models are trained to predict a class label based on the input data features. However, class labels are organized hierar

The Art of Not Forgetting A Local Learning Architecture for Continual Learning

Model ReleasesDGX agent

arXiv:2607.26523v1 Announce Type: new Abstract: We introduce CMP (Cognitive Memory Primitive), a continual-learning architecture that repre?sents inputs as sparse relational codes, stores them in a tw

The Confounder Trap: Treatment-Encoding Representations in Causal Inference with Text

SafetyDGX agent

arXiv:2607.26309v1 Announce Type: cross Abstract: Estimating causal effects of linguistic properties from observational text is difficult because the same document can contain both the treatment of in

The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness

Model ReleasesDGX agent

arXiv:2503.10647v2 Announce Type: replace Abstract: This study evaluated the diagnostic reliability of two Large Language Models (LLMs), Google Gemini 2.0 Flash and OpenAI ChatGPT-4o, across three dim

The Rise of AI in Weather and Climate Information and its Impact on Global Inequality

ApplicationsDGX agent

arXiv:2603.05710v2 Announce Type: replace-cross Abstract: AI development's current trajectory risks automating and amplifying the North-South divide in the global climate information system. Frontier

The Sparsity Ceiling: Where Spiking Networks Can and Cannot Trade Activity for Energy

ResearchDGX agent

arXiv:2607.26648v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) are promoted as an energy-efficient substrate because sparse, event-driven activity replaces dense multiply-accumulates

Think Short, Defer Smart, Act, and Repeat: Calibrated Reasoning and Uncertainty-Aware Deferral for Edge LLM Agents

Local AiDGX agent

arXiv:2607.26865v1 Announce Type: cross Abstract: LLM agents following the ReAct paradigm are promising enablers of complex multi-step tasks, including multi-hop question answering, code generation, a

Thinking Under Uncertainty: Evidence Use and Information-Seeking in Language Models

SafetyDGX agent

arXiv:2607.26845v1 Announce Type: new Abstract: Inference-time thinking improves the performance of large language models, but aggregate outcomes do not reveal whether models use available evidence mo

Tight Generalization Bound for AdaBoost

Model ReleasesDGX agent

arXiv:2607.26838v1 Announce Type: new Abstract: In this paper we show that the generalization error of AdaBoost is Thetaig(frac{dln(ngamma^{2}/d)}{ngamma^2}+frac{ln(1/elta)}{n}ig), where gamma is the

Time-delay Control Using a New Nonlinear Adaptive Law for Cable-Driven Robots

ResearchDGX agent

arXiv:2607.26383v1 Announce Type: cross Abstract: Cable-driven manipulators exhibit strong nonlinearities and low structural stiffness, which make precise control challenging under time-varying uncert

TiPToP: A Modular Open-Vocabulary Robot Manipulation System That Plans

Model ReleasesDGX agent

arXiv:2603.09971v2 Announce Type: replace Abstract: We present TiPToP, a modular manipulation system that integrates pretrained foundation models with a GPU-accelerated Task and Motion Planner to solv

Top-k Pareto Bandits: Hypervolume Regret for Multi-Objective Slate Selection

AgentsDGX agent

arXiv:2607.26273v1 Announce Type: new Abstract: We consider a stochastic multi-objective bandit problem where, at each round, the agent selects a slate of k arms and observes their d-dimensional rewar

Towards Grounded GI Endoscopy VQA via Multi-Task Learning on Small VLMs

SafetyDGX agent

arXiv:2607.27122v1 Announce Type: new Abstract: Gastrointestinal (GI) endoscopic image analysis has shifted from single-label classification toward visual question answering (VQA), where a model must

Towards Trustworthy Embodied Intelligence: A Systems Framework and Graded Trustworthiness Levels

Model ReleasesDGX agent

arXiv:2607.26121v1 Announce Type: new Abstract: Embodied intelligence integrates learned perception and decision making with real-time computation, control, and physical interaction. Because failures

ToxScreen: Detecting Whether an LLM Has Been Poisoned

Model ReleasesDGX agent

arXiv:2607.26849v1 Announce Type: cross Abstract: As large language models (LLMs) are deployed in high-stakes domains, adversaries may poison training data to implant backdoors: hidden triggers that c

TPCD: Tone-Pressure Contrastive Decoding and the Label-Free Gating Bottleneck in Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.26536v1 Announce Type: new Abstract: High-pressure prompts can push vision-language models (VLMs) into unsupported commitments, such as reading illegible text, reporting indeterminate times

TPD: Temporal Prior Decoupling for Text-to-Video Diffusion Models

ResearchDGX agent

arXiv:2607.26706v1 Announce Type: new Abstract: Text-to-video diffusion models generate temporally coherent content from natural language, yet when a prompt describes an early scene that persists whil

← Previous
1…128129130131132…998
Next →