AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
29 Jul 2026

SurgSLOT: Segment Anything in Surgical Videos via Semantic Long-term Tracking

Model ReleasesDGX agent

arXiv:2511.16618v2 Announce Type: replace Abstract: Surgical scene understanding demands temporally consistent tracking of instruments and tissues. For clinical use, such tracking should generalize to

TabRank: Chain-of-Thought Distillation for Table Re-Rankers

Model ReleasesDGX agent

arXiv:2607.25182v1 Announce Type: cross Abstract: The ability to retrieve relevant tables for answering questions is a key task for structured information retrieval. Multi-stage retrieval systems rely

TaylorPODA: A Taylor Expansion-Based Method to Improve Post-Hoc Attributions for Opaque Models

Local AiDGX agent

arXiv:2507.10643v4 Announce Type: replace-cross Abstract: Post-hoc model-agnostic local attribution (LA) methods have been widely adopted to explain opaque AI models by quantifying feature-wise contri

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Temporal-Distance JEPA: Plan-Aware Representation Learning for Latent World Model Predictive Control

TutorialsDGX agent

arXiv:2607.25337v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) learn world models by predicting in representation space rather than reconstructing pixels, making them

The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play

SafetyDGX agent

arXiv:2607.25425v1 Announce Type: new Abstract: Capture the Flag (CTF) competitions are among cybersecurity's most effective training grounds, developing practical skill across cryptography, web explo

The Effect of Text Chunk Size on Retrieval-Augmented Generation Performance

ResearchDGX agent

arXiv:2607.24767v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems have emerged as a powerful process for allowing large language models (LLMs) to retrieve relevant informa

The LAIA Dataset: Labelled Attention for Intelligent Automobiles

AgentsDGX agent

arXiv:2607.25570v1 Announce Type: cross Abstract: The development of autonomous vehicles (AVs) usually relies heavily on data-driven artificial intelligence (AI) models that require large volumes of s

The Possibility of Artificial Intelligence Becoming a Subject and the Alignment Problem

SafetyDGX agent

arXiv:2604.14990v2 Announce Type: replace Abstract: The prospect of Artificial General Intelligence (AGI) is increasingly driving institutional decisions, and alignment of AGI is a hard problem. The c

The Scaling Properties of Implicit Deductive Reasoning in Transformers

SafetyDGX agent

arXiv:2605.04330v2 Announce Type: replace Abstract: We investigate the scaling properties of implicit deductive reasoning over Horn clauses in depth-bounded Transformers. By systematically decorrelati

The User Asks, Platforms Compete: How Agentic Recommendation Markets Take Shape

AgentsDGX agent

arXiv:2607.25253v1 Announce Type: new Abstract: Online recommendation has traditionally taken place after a user enters a platform, which determines the candidate pool and the ranking shown to the use

Three Sides of Retrieval: Factorial Evidence for Document-Side, Query-Side, and Answer-Side Complementarity in RAG

Model ReleasesDGX agent

arXiv:2607.24781v1 Announce Type: cross Abstract: RAG systems rely on chunking, which destroys structural information in documents. Existing heading-based retrieval (Jeong et al., 2025) requires multi

TIGA: Trajectory-Injected Generative Attack against Black-box AIGC Detectors

ResearchDGX agent

arXiv:2607.25894v1 Announce Type: new Abstract: Recent diffusion models have achieved remarkable realism in facial image synthesis, posing growing challenges to artificial intelligence-generated conte

Time-Frequency Consistency Learning for Robust Speech Deepfake Detection

SafetyDGX agent

arXiv:2607.17761v2 Announce Type: replace-cross Abstract: Recently, speech deepfake detection (SDD) has achieved significant progress. However, its robustness evaluation remains largely confined to co

TimeCapsule: Generative Hallucination as a Method for Historical Sensemaking

Model ReleasesDGX agent

arXiv:2607.24750v1 Announce Type: new Abstract: Large Language Models (LLMs) are temporally overexposed: trained on vast contemporary corpora, they encode present-day concepts that make them unreliabl

Tokenizing Numerical and Embedding Features for LLM RecSys

Model ReleasesDGX agent

arXiv:2607.10016v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used as backbone architectures for recommender systems because of their strong sequence modeling

Tokens are All You Need: Dual-purpose Semantic IDs for Achieving LLM-Level I/O Efficiency in recommendation systems

ApplicationsDGX agent

arXiv:2607.24865v1 Announce Type: cross Abstract: Large-scale recommendation systems face 'Memory Wall' bottlenecks due to massive, dense embedding tables. While generative retrieval uses discrete tok

Tools Are Not Islands: Set-Level Tool Retrieval for LLM Agents via Query-Conditioned Hyperedge Prediction

AgentsDGX agent

arXiv:2607.25718v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly rely on invoking external tools to complete real-world tasks. Tool retrieval, which selects a small tas

TopoGR: Revealing and Preserving Latent Structure of Semantic ID in Generative Recommendation

Model ReleasesDGX agent

arXiv:2607.25216v1 Announce Type: cross Abstract: Semantic ID-based generative recommendation tokenizes each item into a sequence of discrete semantic IDs and predicts the next item by generating sema

Toward a systematic method for identifying language areas

ResearchDGX agent

arXiv:2607.25305v1 Announce Type: new Abstract: Macroareas are geographical areas used in typological research for grouping variables of interest. In linguistic typology, languages in a given macroare

Toward an Organizational Science of Multi-Agent LLM Systems: Decoupling Who, How, and Which Algorithm

Model ReleasesDGX agent

arXiv:2607.25446v1 Announce Type: new Abstract: Multi-agent frameworks built on large language models (LLMs) routinely entangle three logically distinct concerns: who is on the team (organization), ho

Toward Standardized Cross-Vendor Agent Tool Trust Management in Autonomous Networks

AgentsDGX agent

arXiv:2607.25914v1 Announce Type: new Abstract: Autonomous Network Levels 4-5 require AI agents to invoke tools across vendor boundaries without human oversight, yet existing management standards lack

Towards a Statistical Understanding of Neural Networks: Beyond the Neural Tangent Kernel Theories

ResearchDGX agent

arXiv:2412.18756v2 Announce Type: replace Abstract: A primary advantage of neural networks lies in their feature learning characteristics, which is challenging to theoretically analyze due to the comp

Towards an Agent Operating System - Lessons from Classical and Cloud OS

AgentsDGX agent

arXiv:2607.25076v1 Announce Type: new Abstract: Every major wave of platform software follows the same arc: an initial period of experimentation with competing frameworks and ad-hoc implementations, f

Towards Embodied Cognition in Robots via Spatially Grounded Synthetic Worlds

HardwareDGX agent

arXiv:2505.14366v2 Announce Type: replace Abstract: We present a conceptual framework for training Vision-Language Models (VLMs) to perform Visual Perspective Taking (VPT), a core capability for embod

Towards Faithful Sentimental Image Captioning via Evidence-Aware Multi-Agent Reasoning

Model ReleasesDGX agent

arXiv:2607.25789v1 Announce Type: new Abstract: Sentimental Image Captioning (SIC) requires balancing emotional expression with visual fidelity. Existing methods often struggle with this trade-off, le

Towards real-time surrogate-free Bayesian inversion for neutron reflectometry

ResearchDGX agent

arXiv:2509.06924v3 Announce Type: replace Abstract: Neutron reflectometry (NR) is a key enabling technology for many areas of scientific development. Although the forward reflectivity model is well-kn

Towards Real-World RUL Prediction for Aircraft Compressors Under Variable Conditions with Physics-Guided Domain Adaptation

Model ReleasesDGX agent

arXiv:2607.24809v1 Announce Type: cross Abstract: Remaining useful life prediction for aircraft centrifugal air compressors in real commercial operations poses challenges that controlled benchmark dat

Towards Reliable Stain Transfer: An Iterative Data-Model Co-Optimization Framework Based on Multimodal Expert-Guided Assessment

ResearchDGX agent

arXiv:2607.25393v1 Announce Type: new Abstract: Histopathological examination primarily relies on hematoxylin and eosin (H&E) and immunohistochemistry (IHC) staining. Although IHC provides critical mo

Towards Robust Reinforcement Learning for Small-Scale Language Model Agents

Model ReleasesDGX agent

arXiv:2607.25091v1 Announce Type: new Abstract: The alignment of Small Language Models (SLMs) in the 70--500M parameter range using reinforcement learning is often considered unstable, though the unde

Towards Understanding the Cognitive Habits of Large Reasoning Models

Model ReleasesDGX agent

arXiv:2506.21571v3 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs), which autonomously produce a reasoning Chain of Thought (CoT) before producing final responses, offer a promisi

Track-Leakage-Free Hold-Out Self-Validation for Photogrammetric Reconstruction: Protocol, Sensitivity, and Limits

Model ReleasesDGX agent

arXiv:2607.24852v1 Announce Type: new Abstract: Automated photogrammetric inspection emits metric measurements from a 3D reconstruction whose own correctness is normally unknown without an external su

Transformer Transformer: A Unified Model for Motion-Conditioned Robot Co-design

ResearchDGX agent

arXiv:2607.25798v1 Announce Type: new Abstract: An often overlooked factor of robot manipulation performance is the embodiment of the robot itself. Motivated by this problem, we study motion-condition

Tri-Manual Visuomotor Imitation Learning of Robot Policies

SafetyDGX agent

arXiv:2607.25731v1 Announce Type: new Abstract: Bimanual teleoperation provides an effective way to collect robot demonstrations, but it assumes that the operator and robot have matching numbers of si

Tripody: An Overconstrained 3-SPR-like Parallel Robot for High-Reach Construction Tasks

SafetyDGX agent

arXiv:2607.25781v1 Announce Type: new Abstract: Many ceiling construction tasks still rely on heavy serial manipulators that are difficult to deploy in cluttered interiors, motivating lightweight, fie

TRWH: A Text-Driven Random Walk Heterogeneous GNN for Semantic-Aware Sparse Recommendation

ResearchDGX agent

arXiv:2607.25471v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) and Large Language Models (LLMs) have each advanced recommendation systems by modeling structural and semantic signals, res

TWICE: Two-Clock, Two-Window Learning for Long-Horizon Conversion Prediction in Online Advertising

Model ReleasesDGX agent

arXiv:2607.25404v1 Announce Type: new Abstract: Long-horizon conversion prediction under delayed feedback creates a two-clock, two-window learning problem in online advertising. A short base observati

Two Views, One Voice: Evidence-Grounded Conversational Music Recommendation

Model ReleasesDGX agent

arXiv:2607.24846v1 Announce Type: cross Abstract: Traditional conversational recommenders entangle retrieval and response generation within a single text interface, so exact entity cues fade as the di

Understanding Semantic IDs: From Item Representation to Item Selection in Generative Recommendation

Local AiDGX agent

arXiv:2607.24995v1 Announce Type: new Abstract: Semantic IDs (SIDs) are now a central component of generative recommendation. Current SID-based systems assign three roles to the same token sequence. S

Understanding User Experiences of Computer Use Agents: Design Space and Opportunities for Building Agent UX Prototypes

AgentsDGX agent

arXiv:2510.04452v3 Announce Type: replace-cross Abstract: Computer use agents (or 'agents') are generative AI that automates actions within user interfaces from user commands. Current research focuses

Unified Semantic Modeling Framework for Large-Scale Job Understanding at LinkedIn

ResearchDGX agent

arXiv:2607.24783v1 Announce Type: new Abstract: Job understanding is critical to LinkedIn's mission of connecting talent with opportunity. This task involves transforming unstructured and noisy job po

Unifying Active Learning and Semi-Supervised Learning for Medical Image Segmentation

TutorialsDGX agent

arXiv:2607.25014v1 Announce Type: new Abstract: In practical settings, medical image segmentation models are often developed with limited annotated data rather than fully labeled datasets. Training fr

UniMem: Complementary Episodic-to-Parametric Memory for Boundary-Agnostic Task Streams

Model ReleasesDGX agent

arXiv:2607.26017v1 Announce Type: new Abstract: Memory is essential for LLM agents to accumulate task experience and reuse task-specific execution strategies. However, real-world deployment over bound

Universal Pansharpening Model

Model ReleasesDGX agent

arXiv:2603.03831v2 Announce Type: replace Abstract: Pansharpening generates the high-resolution multi-spectral (MS) image by integrating spatial details from a texture-rich panchromatic (PAN) image an

Unlocking Spatial Grounding in Large Audio-Visual Retrieval models

SafetyDGX agent

arXiv:2607.24786v1 Announce Type: cross Abstract: Weak supervision sets a practical regime for audio-visual sound source localization as dense spatial annotations are costly to obtain at scale. The ta

Untangling Co-Drift: Proactive Multi-Intent Failure Prediction and Root-Cause Disambiguation for Self-Driving Networks

Model ReleasesDGX agent

arXiv:2607.25989v1 Announce Type: cross Abstract: The vision of self-driving networks that monitor, reason, and act upon themselves with minimal human intervention relies on tightly coupled monitoring

Untrusted Authors, Trusted Answers: A Calculus of Fidelity-Graded Translations

ResearchDGX agent

arXiv:2607.14137v2 Announce Type: cross Abstract: To answer a question about a program, move the program to where the question is decidable. Every such move is a translation, and every translation is

Using Data-Derived Priors to Guide CNN Architecture Design for NIR Chemometrics

TutorialsDGX agent

arXiv:2607.25636v1 Announce Type: new Abstract: Convolutional neural networks (CNN) for near-infrared (NIR) chemometrics are often designed using generic architectural rules, although spectral dataset

VAD to the Bone: Ultra-Tiny Speech Activity Detection for Edge Deployment

ResearchDGX agent

arXiv:2607.25870v1 Announce Type: cross Abstract: Voice activity detection (VAD) triggers downstream speech processing in always-on systems under strict memory, latency, and compute constraints. Recen

VaLiDRec: Variable-Length LLM-Aligned Semantic IDs for Generative Recommendation

ApplicationsDGX agent

arXiv:2607.25209v1 Announce Type: cross Abstract: Generative recommendation commonly represents items using fixed-length semantic identifiers (SIDs) constructed through clustering and quantization. Ho

Variance-Reduced Conditional Gradient Methods under Markovian Sampling for Nonconvex Composite Optimization

SafetyDGX agent

arXiv:2607.25785v1 Announce Type: cross Abstract: We study stochastic composite nonconvex optimization over a compact convex set when gradient samples arrive along a single trajectory of a fixed ergod

Vector-Valued Distributional Reinforcement Learning Policy Evaluation: A Hilbert Space Embedding Approach

SafetyDGX agent

arXiv:2601.18952v2 Announce Type: replace Abstract: We propose an (offline) multi-dimensional distributional reinforcement learning framework (KE-DRL) that leverages Hilbert space mappings to estimate

Verification Without Distrust: Reframing User-Side Oversight as Routine Epistemic Governance in Everyday Human-Chatbot Interaction

ResearchDGX agent

arXiv:2607.24761v1 Announce Type: cross Abstract: Research on human-AI interaction has long framed verification of system outputs as a trust-contingent behavior that better-calibrated trust should red

VetClaw: An Edge-Cloud Multimodal Agentic System for Veterinary Disease Screening

SafetyDGX agent

arXiv:2607.26042v1 Announce Type: new Abstract: We present VetClaw, an edge-cloud multimodal agentic system for early veterinary disease screening. VetClaw uses a camera module as an edge sensing devi

VisRAG2.0: Mitigating Visual Hallucinations via Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2510.09733v2 Announce Type: replace Abstract: Visual Retrieval-Augmented Generation (VRAG) has emerged as a promising paradigm for equipping Vision-Language Models (VLMs) with external visual ev

Visual prompt engineering for video models

ResearchDGX agent

arXiv:2607.25537v1 Announce Type: cross Abstract: In the age of foundation models, a model is only as good as its prompt. For this reason, prompt engineering has become an essential technique for impr

VisualPatchWorld: Code World Models as Latent Structured Representations for Planning

TutorialsDGX agent

arXiv:2607.25236v1 Announce Type: new Abstract: Different research lines use the term world model in different ways, yet they share a common aim: to capture how the world evolves under action in a for

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents

AgentsDGX agent

arXiv:2607.24748v1 Announce Type: cross Abstract: Visually-rich documents such as reports, slides, and manuals often distribute the evidence needed to answer a question across multiple pages, mixing t

Wall Shear Stress Reconstruction from Concentration: Differentiable Physics and Physics-Informed Neural Networks

Model ReleasesDGX agent

arXiv:2606.06313v2 Announce Type: replace-cross Abstract: Wall shear stress (WSS) governs near-wall transport dynamics and is a key hemodynamic indicator in cardiovascular flows, yet remains difficult

WALoMA: A Multitask Wireless Foundation Model via Adaptive Low-Rank Masked Autoencoders

Model ReleasesDGX agent

arXiv:2607.25763v1 Announce Type: cross Abstract: This paper proposes a multitask wireless foundation model via adaptive low-rank masked autoencoders (WALoMA), a unified multi-task foundation model fo

'We'll have to see how it works': An interview study to understand collaborative practices in interdisciplinary artificial intelligence and healthcare research

ApplicationsDGX agent

arXiv:2311.18424v3 Announce Type: replace-cross Abstract: Developing artificial intelligence (AI) algorithms for healthcare is a collaborative effort, bringing data scientists, clinicians, patients an

← Previous
1…138139140141142…998
Next →