AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,405 results
8 Jun 2026

Beyond Backscatter: InSAR coherence from detected SAR images

TutorialsDGX agent

arXiv:2606.07374v1 Announce Type: cross Abstract: In this work, we propose a deep learning framework for coherence regression directly from detected SAR images, without the need for accurate coregistr

Beyond Linear and Overcomplete Regimes: A Mean-Field Analysis of Bottleneck Autoencoders

TutorialsDGX agent

arXiv:2606.07120v1 Announce Type: new Abstract: Autoencoders (AEs) learn low-dimensional representations by mapping data into a latent space while minimizing reconstruction error. Despite their empiri

Broadband Hyperspectral 3D Imaging using Dispersed Structured Light

HardwareDGX agent

arXiv:2605.25757v2 Announce Type: replace Abstract: Hyperspectral 3D imaging enables the capture of dense spectral information and scene geometry but has traditionally been confined to narrow spectral

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

CAPE: Contrastive Action-conditioned Parallel Encoding for Embodied Planning

TutorialsDGX agent

arXiv:2606.07304v1 Announce Type: new Abstract: Embodied agents need to predict the future consequences of candidate actions in order to plan effectively before execution. Existing visual dynamics mod

CARVE-Q: Quantum-Proposed, Classically Certified Interactive Driving Repair

SafetyDGX agent

arXiv:2606.06531v1 Announce Type: new Abstract: The critical question after a correct driving veto is not only whether a maneuver is unsafe, but whether the blocked interaction admits a lawful, audita

Characterization of Gaussian Universality Breakdown in High-Dimensional Empirical Risk Minimization

ResearchDGX agent

arXiv:2604.03146v2 Announce Type: replace-cross Abstract: We study high-dimensional convex empirical risk minimization (ERM) under general non-Gaussian data designs. By heuristically extending the Con

ChemQuests: A Curated Chemistry Question-Answer Database Extracted from ChemRxiv papers

ApplicationsDGX agent

arXiv:2505.05232v3 Announce Type: replace Abstract: The rapid expansion of chemistry literature poses significant challenges for researchers seeking to efficiently access domain-specific knowledge. To

COMPOSE: Hypergraph Cover Optimization for Multi-view 3D Human Pose Estimation

ResearchDGX agent

arXiv:2601.09698v2 Announce Type: replace Abstract: 3D human pose estimation from sparse multi-view camera rigs is an essential task for numerous applications, including action recognition, sports ana

Conformal Disentanglement and Latent-Space Curation: A Neural Framework for Perspective Synthesis, Differentiation and Targeted Generation

ResearchDGX agent

arXiv:2408.15344v2 Announce Type: replace Abstract: Many scientific and engineering problems involve observing a common phenomenon through multiple heterogeneous sensors or measurement modalities. Suc

Constructing VAE Latent Spaces with Prescribed Topology

TutorialsDGX agent

arXiv:2606.07058v1 Announce Type: cross Abstract: Variational autoencoders (VAEs) learn low-dimensional latent representations of high-dimensional data. When the data lies on a manifold with non-Eucli

Contrastive Training with LLM-generated Near-Misses for Robust Code-Switching Speech Recognition

ResearchDGX agent

arXiv:2606.06985v1 Announce Type: new Abstract: Code-switching (CS), the alternation between multiple languages within a single utterance, remains challenging for Automatic Speech Recognition (ASR). T

Deep Agents explained in <90 seconds by @sydneyrunkle

AgentsDGX agent

Deep Agents are an AI concept that Sydney Runkle explains concisely in under 90 seconds, likely covering how agents can be designed to operate with deeper reasoning and decision-making capabilities. T

Discovering Interpretable Algorithms by Decompiling Transformers to RASP

ResearchDGX agent

arXiv:2602.08857v2 Announce Type: replace-cross Abstract: Recent work has shown that the computations of Transformers can be simulated in the RASP family of programming languages. These findings have

Does Appearance Help? A Systematic Study of Image-Based Re-Identification in Online 3D Multi-Pedestrian Tracking

ResearchDGX agent

arXiv:2606.07233v1 Announce Type: new Abstract: LiDAR-based 3D Multi-Object Tracking (MOT) typically relies solely on geometric information, which is often insufficient to distinguish between targets

DSU-Net: An Attention-Enhanced Dense Skip U-Net for Breast Lesion Segmentation in Mammographic Images

ResearchDGX agent

arXiv:2606.06537v1 Announce Type: cross Abstract: Breast cancer remains one of the leading causes of cancer-related mortality among women worldwide, making early detection essential for effective trea

Dual Latent Memory for Visual Multi-agent System

AgentsDGX agent

arXiv:2602.00471v2 Announce Type: replace Abstract: While Visual Multi-Agent Systems (VMAS) promise to enhance comprehensive abilities through inter-agent collaboration, empirical evidence reveals a c

End-to-end encrypted ML inference with Amazon SageMaker AI and FHE

TutorialsDGX agent

This blog has previously discussed FHE for ML inference in the post Enable fully homomorphic encryption with Amazon SageMaker endpoints for secure, real-time inferencing, but this post goes a little f

Evaluate your Amazon Nova Sonic voice agent at scale, no microphone required

AgentsDGX agent

In this post, we walk you through the Nova Sonic Test Harness, an open source framework that we built to solve both problems. It serves as a rapid iteration tool for tuning system prompts and tool con

Feasible Action Space Reduction for Quantifying Causal Responsibility in Continuous Spatial Interactions

AgentsDGX agent

arXiv:2505.17739v2 Announce Type: replace-cross Abstract: Understanding the causal influence of one agent on another agent is crucial for safely deploying artificially intelligent systems such as auto

Geometric-Aware Hypergraph Reasoning for Novel Class Discovery in Point Cloud Segmentation

ResearchDGX agent

arXiv:2606.07280v1 Announce Type: new Abstract: Novel class discovery in point cloud segmentation aims to transfer knowledge from known classes to automatically identify and segment unlabeled novel cl

Geometric Second-Order Feature Correlation Learning for Self-Supervised Speech Emotion Recognition

ResearchDGX agent

arXiv:2606.06550v1 Announce Type: cross Abstract: Self-supervised learning (SSL) yields powerful, context-rich representations for speech emotion recognition (SER), yet aggregating these representatio

GOPAgen: Motion-Aware and Efficient Agentic Long-Video Understanding with Structural Memory and Hierarchical Reasoning

AgentsDGX agent

arXiv:2606.06532v1 Announce Type: new Abstract: Despite significant progress in agentic long video understanding, existing methods still lack detailed motion comprehension coupled with an efficient me

Heterogeneous Effects of Green Finance on Urban Decarbonization: Evidence from 285 Cities in China

SafetyDGX agent

arXiv:2606.06986v1 Announce Type: new Abstract: While green finance has become a key instrument for low-carbon city transitions, its actual decarbonization effects and transmission mechanisms remain u

Hi everyone, I'm building a local AI agent with Ollama and exploring dynamic PDF extraction. Since Ollama can't directly process PDFs, I'm extracting text and passing it via prompts. Should I use PDFPlumber, a vector database (RAG), or another approach for accurate document understanding? Guide me !

Local AiDGX agent

User seeks guidance on PDF processing methods for local AI agents built with Ollama, comparing approaches like PDFPlumber extraction, vector database RAG systems, and alternative techniques for accura

HORUS: A Mixed Reality Interface for Managing Teams of Mobile Robots

ResearchDGX agent

arXiv:2506.02622v2 Announce Type: replace Abstract: Mixed Reality (MR) interfaces have been extensively explored for controlling mobile robots, but there is limited research on their application to ma

i expect value for money to increase over time as competitive pressures force providers to lower prices. but in the short time, looks like t…

SafetyDGX agent

i expect value for money to increase over time as competitive pressures force providers to lower prices. but in the short time, looks like the opposite may be happening: Does a token buy you more or l

Inside the Visual Mind: Neuroscience-Motivated Concept Circuits for Interpreting and Steering Vision Transformers

ResearchDGX agent

arXiv:2606.06664v1 Announce Type: cross Abstract: Despite high accuracy, Vision Transformer (ViT) predictions can be driven by spurious cues, raising the need to understand their inner workings before

KIT's Submission to Cross-Lingual Voice Cloning in IWSLT 2026

ResearchDGX agent

arXiv:2606.07240v1 Announce Type: new Abstract: Cross-lingual voice cloning aims to generate speech in a target language while preserving speaker identity from a source-language reference. This task i

Learning All-Terrain Locomotion for a Planetary Rover with Actively Articulated Suspension

SafetyDGX agent

arXiv:2606.06790v1 Announce Type: cross Abstract: This paper presents ERNEST, a four-wheeled planetary rover concept equipped with a two-degree-of-freedom Active Gimbal Suspension that combines yaw an

Learning to Execute Graph Algorithms Exactly with Graph Neural Networks

Local AiDGX agent

arXiv:2601.23207v2 Announce Type: replace-cross Abstract: Understanding what graph neural networks can learn, especially their ability to learn to execute algorithms, remains a central theoretical cha

LLM-Augmented Digital Twin for Policy Evaluation in Short-Video Platforms

SafetyDGX agent

arXiv:2603.11333v2 Announce Type: replace Abstract: Short-video platforms are closed-loop, human-in-the-loop ecosystems where platform policy, creator incentives, and user behavior co-evolve. This fee

LLM-Guided Search for Deletion-Correcting Codes

ResearchDGX agent

arXiv:2504.00613v2 Announce Type: replace Abstract: Finding deletion-correcting codes of maximum size has been an open problem for over 70 years, even for a single deletion. We adapt FunSearch, a larg

London-based PhysicsX, which uses AI to design industrial parts like jet engines and semiconductors, raised a 300M Series C led by Temasek at a 2.4B valuation (Mark Bergen/Bloomberg)

ApplicationsDGX agent

Mark Bergen / Bloomberg: London-based PhysicsX, which uses AI to design industrial parts like jet engines and semiconductors, raised a 300M Series C led by Temasek at a 2.4B valuation — PhysicsX, a Br

LRMIL: Efficient Low-Resolution Multiple Instance Learning via High-Resolution Knowledge Distillation for Whole Slide Image Classification

ApplicationsDGX agent

arXiv:2606.06864v1 Announce Type: new Abstract: Multiple instance learning (MIL) has become a standard paradigm for whole slide image (WSI) analysis in digital pathology, as it enables slide-level pre

Meaning in Order, Order in Meaning: Semantic R-precision for Keyphrase Evaluation

ResearchDGX agent

arXiv:2606.07057v1 Announce Type: cross Abstract: Evaluating the quality of automatically generated keyphrases remains a complex challenge. Traditional metrics either rely on exact lexical matching or

Measuring Agents in Production

AgentsDGX agent

arXiv:2512.04123v4 Announce Type: replace-cross Abstract: LLM-based agents already operate in production across many industries, yet we lack an understanding of what technical methods make deployments

Mind the Gap: Disentangling Performance Bottlenecks in Video Instance Segmentation

ResearchDGX agent

arXiv:2606.07394v1 Announce Type: new Abstract: In Video Instance Segmentation (VIS), classification, segmentation, and tracking objectives are jointly evaluated, but their individual contributions to

Network Recovery from Cascade Data: A Debiased Jacobian-Based Machine Learning Approach

SafetyDGX agent

arXiv:2606.07483v1 Announce Type: new Abstract: Many important outcomes unfold as dynamic cascades, including product adoption, disease spread, financial distress, and information diffusion. A central

Neuro-Symbolic Learning for Long-Horizon Task Planning Under Complex Logical Constraints

SafetyDGX agent

arXiv:2606.06877v1 Announce Type: cross Abstract: Task planning often suffers from severe efficiency bottlenecks when robots must reason over long-horizon action sequences under complex logical constr

New paper on how AI agents are reshaping knowledge work. This is a nice economic read on where agents actually change knowledge work to meet…

AgentsDGX agent

New paper on how AI agents are reshaping knowledge work. This is a nice economic read on where agents actually change knowledge work to meet that gap directly. (bookmark it) It studies agent adoption

Off-Policy Evaluation with Strategic Agents via Local Disclosure

Local AiDGX agent

arXiv:2606.07308v1 Announce Type: new Abstract: We study off-policy evaluation (OPE) under strategic behavior where decision subjects (or agents) respond to a decision maker's policy by strategically

OGA-AID: Clinician-in-the-loop AI Report Drafting Assistant for Multimodal Observational Gait Analysis in Post-Stroke Rehabilitation

AgentsDGX agent

arXiv:2604.05360v2 Announce Type: replace-cross Abstract: Gait analysis is essential in post-stroke rehabilitation but remains time-intensive and cognitively demanding, especially when clinicians must

On theCUBE Pod: Snowflake, Cisco and the race to control the AI stack

IndustryDGX agent

Artificial intelligence companies are competing for dominance over the emerging AI stack, and it’s anyone’s game. The AI market is maturing. SpaceX Corp. wants to raise $75 billion to go public, and A

Optimal Control Approach for Non-prehensile Ball Juggling Using a 7-DoF Manipulator

ApplicationsDGX agent

arXiv:2606.06704v1 Announce Type: new Abstract: Non-prehensile object manipulation skills are important for real-world robot interactions, enabling highly dynamic tasks such as balancing a glass on a

Predictable Compression Failures: Order Sensitivity and Information Budgeting for Evidence-Grounded Binary Adjudication

ResearchDGX agent

arXiv:2509.11208v3 Announce Type: replace-cross Abstract: Transformers used for evidence-grounded binary adjudication (e.g., support/refute, yes/no, or verifier-backed pass/fail decisions) can be sens

Predicting Dynamic Map States from Limited Field-of-View Sensor Data

AgentsDGX agent

arXiv:2602.12360v2 Announce Type: replace Abstract: When autonomous systems are deployed in real-world scenarios, sensors are often subject to limited field-of-view (FOV) constraints, either naturally

Probabilistic learning to perform pre-onset individualised prediction of disease severity: application to Veno Occlusive Disease

ResearchDGX agent

arXiv:2606.06516v1 Announce Type: cross Abstract: We advance a new probabilistic supervised learning approach that permits reliable, automated, and early individualised prediction of the severity with

Quantifying Media Representation Dynamics Across 25 Years of News Reporting on Policing-related Deaths

ResearchDGX agent

arXiv:2606.06812v1 Announce Type: new Abstract: We perform the largest known computational analysis of Canadian news narratives about police-involved deaths, spanning 4,000 articles from the last quar

Reactivity-Informed Machine Learning for Performance Prediction and Design Space Exploration of Alkali-Activated Slag

TutorialsDGX agent

arXiv:2606.06765v1 Announce Type: cross Abstract: Establishing quantitative relationships among mix design, raw material properties, curing conditions, and performance remains a long-standing challeng

Semantic-Structural Alignment for Generative Pictorial Charts

SafetyDGX agent

arXiv:2606.06498v1 Announce Type: cross Abstract: Traditional statistical graphics are precise but often lack the visual appeal, memorability, and engagement of pictorial charts. We present a generati

Signal-Driven Observation for Long-Horizon Web Agents

AgentsDGX agent

arXiv:2606.06708v1 Announce Type: new Abstract: Web agents operating over long horizons ingest raw DOM and accessibility trees -- routinely tens of thousands of tokens -- at every action step, causing

Standard vs. Modular Sampling: Best Practices for Reliable LLM Unlearning

ApplicationsDGX agent

arXiv:2509.05316v2 Announce Type: replace-cross Abstract: A conventional LLM Unlearning setting consists of two subsets -'forget' and 'retain', with the objectives of removing the undesired knowledge

TargetSEC: Plug-and-Play In-the-Wild Speech Emotion Conversion via Arousal-Conditioned Latent Style Diffusion

ApplicationsDGX agent

arXiv:2606.07293v1 Announce Type: cross Abstract: Speech Emotion Conversion (SEC) aims to transform the emotion of a source utterance into a target emotion while preserving content and speaker identit

Telling stories, making Hanzi: AI-assisted co-creation with elderly migrants in urban China

ResearchDGX agent

arXiv:2507.01548v3 Announce Type: replace-cross Abstract: This paper explores how older migrants in urban China can record stories that everyday language and design often miss. We ran two co-creation

The Capacity of Information-Theoretic Secure Aggregation in Federated Learning

Local AiDGX agent

arXiv:2606.07277v1 Announce Type: cross Abstract: Secure aggregation allows a server to aggregate users' local updates while preserving update privacy. Existing information-theoretic problems typicall

Three-dimensional hydro-cluttered locomotion by an undulatory robot

ResearchDGX agent

arXiv:2606.06829v1 Announce Type: new Abstract: Aquatic robots have expanded human access to underwater environments, yet many underwater spaces contain obstacles that can disrupt open-water locomotio

Towards Unified Song Generation and Singing Voice Conversion with Accompaniment Co-Generation

TutorialsDGX agent

arXiv:2606.07015v1 Announce Type: cross Abstract: While song generation and singing voice conversion (SVC) have evolved significantly, they have long been developed isolated: the former lacks zero-sho

Training for Technology: Adoption and Productive Use of Generative AI in Legal Analysis

ApplicationsDGX agent

arXiv:2603.04982v3 Announce Type: replace-cross Abstract: Can targeted user training unlock the productive potential of generative artificial intelligence in professional settings? We study this quest

TraRA: Trajectory-level Recognition Aggregation for Video Text Spotting in Urban Surveillance

ResearchDGX agent

arXiv:2606.07161v1 Announce Type: new Abstract: Video Text Spotting (VTS) is essential for urban surveillance and intelligent transportation systems, enabling automated reading of street signs, vehicl

TrioPose: Native Triple-Stream Diffusion Transformers for Pose-Guided Text-to-Image Generation

SafetyDGX agent

arXiv:2606.07053v1 Announce Type: new Abstract: Pose-guided text-to-image generation often suffers from limb distortions and feature crosstalk in complex multi-person scenarios. While existing UNet-ba

← Previous
1…928929930931932…1007
Next →