AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
Human
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,019 results
9 Jun 2026

Diverse Thinking Schemata Elicit Better Reasoning in Large Language Models

SafetyDGX agent

arXiv:2606.08974v1 Announce Type: new Abstract: Large reasoning models (LRMs) have attracted increasing attention for their ability to solve complex mathematical problems by generating extended reason

DIYHealth Suite: Dataset, Model, and Benchmark for Health Management at Home

Model ReleasesDGX agent

arXiv:2606.07542v1 Announce Type: cross Abstract: Generative AI is reshaping healthcare, yet most existing advances rely on hospital-grade devices, which limits their accessibility and potential for h

DN-Hypo-Pipeline: An AI-Driven Workflow for Hypothesis Generation via Large Language Models and Scientific Explanations

ResearchDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.08532v1 Announce Type: new Abstract: A scientific hypothesis is the first step in research and undergoes experimental validation, yet it also reflects a deep understanding of and reasoning

Do Video Foundation Models Understand Intuitive Physics? A Layerwise Probing Analysis

ResearchDGX agent

arXiv:2606.09646v1 Announce Type: cross Abstract: We study whether pretrained video foundation models encode intuitive-physics information in their frozen representations, and how this information var

Do VLMs See What Sensors Feel? A Scalable Expert-Guided Design for Wheelchair Accessibility Assessment from Street View

SafetyDGX agent

arXiv:2606.07642v1 Announce Type: new Abstract: Assessing built-environment interaction, such as wheelchair accessibility, is difficult because real-world mobility is shaped by distributed, context-de

Docs: ~34K Instagram accounts, including Obama's White House account, were affected in the breach tied to Meta's AI chatbot; attackers changed 3,500+ usernames (New York Times)

IndustryDGX agent

New York Times: Docs: ~34K Instagram accounts, including Obama's White House account, were affected in the breach tied to Meta's AI chatbot; attackers changed 3,500+ usernames — The flaw, which Meta s

Does Persona Make LLMs K-pop Fans? A Pilot Study of LLM-Based Online Concert Audience Agents

SafetyDGX agent

arXiv:2606.07837v1 Announce Type: cross Abstract: A concert is a collective experience, but recorded performance videos are typically watched alone, stripping away the shared audience presence that ma

DOG-DPO:Dynamic Optimization in Geometry for Safety Alignment

SafetyDGX agent

arXiv:2606.07678v1 Announce Type: cross Abstract: Safety alignment for large language models relies on preference data, but current pipelines often train on large, redundant datasets. Existing data se

Domain-Adapted Small Language Models with Hybrid Post-Processing: Achieving Cost-Efficient, Low-Latency Multi-Label Structured Prediction via LoRA Fine-Tuning on Scarce Data

Model ReleasesDGX agent

arXiv:2606.05781v2 Announce Type: replace Abstract: Deploying frontier large language models (LLMs) for domain-specific structured evaluation tasks incurs prohibitive latency, cost, and data-privacy o

Domain Search is now available through the Vercel CLI

ToolsDGX agent

Vercel has added domain search functionality to its command-line interface (CLI), allowing developers to search for and potentially register domains directly from the terminal. This feature integrates

DOME: Learning Transferable Domain Variables from Sparse Supervision for Test-Time Adaptation

ApplicationsDGX agent

arXiv:2606.07646v1 Announce Type: cross Abstract: Test-time adaptation (TTA) aims to align a model to shifting test domains using only unlabeled streaming data. Most existing methods implicitly infer

Don’t always agree with @mustafasuleyman but in this case he is right to call out @AnthropicAI.

Model ReleasesDGX agent

Don’t always agree with @mustafasuleyman but in this case he is right to call out @AnthropicAI. Microsoft AI head calls out Anthropic for acting like Claude is conscious https://www.theverge.com/tech/

@dpetrou @karpathy Yes. Locking in a permanent status quo power structure. Incredibly unsafe, and damaging for humanity's prospects.

SafetyDGX agent

Gary Marcus argues that establishing a permanent, locked-in power structure is fundamentally unsafe and harmful to humanity's long-term prospects. The statement appears to be part of a discussion with

Dr. SHAP-AV: Decoding Relative Modality Contributions via Shapley Attribution in Audio-Visual Speech Recognition

SafetyDGX agent

arXiv:2603.12046v2 Announce Type: replace-cross Abstract: Audio-Visual Speech Recognition (AVSR) leverages both acoustic and visual information for robust recognition under noise. However, how models

Dream-Tac: A Unified Tactile World Action Model for Contact-Rich Robot Manipulation

SafetyDGX agent

arXiv:2606.08737v1 Announce Type: new Abstract: World action models inherit the predictive capability of world models, enabling action generation to be guided by anticipated future observations. Howev

DriveReward: A Comprehensive Dataset and Generative Vision-Language Reward Model for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.08525v1 Announce Type: new Abstract: Reward models play a pivotal role in reinforcement learning (RL) and multi-modal trajectory selection for autonomous driving. However, acquiring such re

Driving Video Retrieval for Complex Queries with Structured Grounding

Model ReleasesDGX agent

arXiv:2606.09109v1 Announce Type: new Abstract: Video retrieval at scale is central to data curation and safety validation in autonomous driving, where users want to find not only scenes but also dyna

Drone boat picked up downed US Army helicopter pilots—a first for sea rescues

IndustryDGX agent

A US military autonomous surface vessel marked the first publicized use of an unmanned drone boat to locate and retrieve downed aircrew in real-world warfare. The unmanned vessel helped rescue two Arm

DroneDAR: Long-Range Drone Distance Estimation Using Monocular Vision and Bounding-Box Features

ApplicationsDGX agent

arXiv:2606.07756v1 Announce Type: new Abstract: Accurate distance estimation for small drones in long-range imagery is important for tracking and situational awareness, yet remains challenging due to

DSFNet: Learning Dual-Domain Spectral Operators for Multi-Modality Spatio-Temporal Forecasting in Urban Transportation Systems

ApplicationsDGX agent

arXiv:2606.07695v1 Announce Type: cross Abstract: Multi-Modality Spatio-Temporal Forecasting (MoSTF) extends traditional spatio-temporal forecasting by incorporating diverse traffic modalities. Despit

DTEX adds AI Risk Management to track how agents and employees use AI

AgentsDGX agent

Behavioral intelligence security company DTEX Systems Inc. today introduced an expanded AI Risk Management product that reads the intent behind how employees and autonomous artificial intelligence age

Dual Quaternion-Based Unscented Kalman Filter with Visual Inertial Odometry for Navigation in GPS-Denied Environments

Model ReleasesDGX agent

arXiv:2606.09292v1 Announce Type: new Abstract: Reliable navigation in GPS-denied environments remains a fundamental challenge in robotics, aerospace, and autonomous vehicle applications. This paper p

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning

SafetyDGX agent

arXiv:2606.08035v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a leading paradigm for enhancing visual reasoning in Multimodal Large Language Mode

DynaCF: Mitigating Shortcut Learning in Reward Models via Dynamic Counterfactual Sensitivity

ResearchDGX agent

arXiv:2606.09043v1 Announce Type: new Abstract: Reward models trained from pairwise preferences often exploit superficial shortcut cues rather than learning true response quality. We propose DynaCF, a

Dynamic Distributed Constraint Optimization and Metareasoning for Continual, Large-Scale Satellite Operations

Local AiDGX agent

arXiv:2601.06188v3 Announce Type: replace Abstract: As Earth-observing satellite constellations grow in size and capability, distributed onboard control offers a pathway to novel responses and time-se

DynaOD: Dynamic Origin-Destination Flow Generation with Discrete-to-Continuous Temporal Semantic Modeling

ApplicationsDGX agent

arXiv:2606.09086v1 Announce Type: new Abstract: Dynamic origin-destination (OD) flow generation seeks to synthesize realistic mobility dynamics from temporal context alone, without relying on historic

Earlytrade raises $10M to bring agentic AI to construction payments

AgentsDGX agent

Earlytrade Pty. Ltd., a company solving payments flow for contractors in the construction industry, today announced it has raised about 10 million in new funding. Today’s capital infusion brings the t

eBook: Exploring the Future of AI Infrastructure

ApplicationsDGX agent

As AI moves from experimentation to production, infrastructure requirements are evolving rapidly. This eBook examines the market, technology, and operational shifts driving the next generation of clou

Echo-DM: Ultrasound Marker Removal via Conditional Latent Diffusion and Region-Aware Fusion

Model ReleasesDGX agent

arXiv:2606.09378v1 Announce Type: new Abstract: Clinical ultrasound images often contain artificial markers, such as measurement calipers and text, to assist diagnostic interpretation and comparison.

Echo-Memory: A Controlled Study of Memory in Action World Models

ResearchDGX agent

arXiv:2606.09803v1 Announce Type: new Abstract: We present extbf{Echo-Memory}, a controlled study of memory mechanisms in action-conditioned world models. These models generate multi-segment videos fr

Edge-Constrained UAV Small-Object Detection with P2 Enhancement and Quantum-Inspired Lightweight Structure Search

ResearchDGX agent

arXiv:2606.09081v1 Announce Type: new Abstract: Unmanned aerial vehicle (UAV) object detection requires compact detectors that retain small-object details under onboard computation and memory constrai

Edge Markets, which provides banking tools for gambling and prediction markets, including real-time payments on Kalshi, raised a $29.2M Series A led by CoinFund (Davis Giangiulio/CNBC)

IndustryDGX agent

Davis Giangiulio / CNBC: Edge Markets, which provides banking tools for gambling and prediction markets, including real-time payments on Kalshi, raised a $29.2M Series A led by CoinFund — As predictio

EditSR: Enhancing Neural Symbolic Regression via Edit-based Rectification

TutorialsDGX agent

arXiv:2606.07915v1 Announce Type: new Abstract: Neural symbolic regression models improve inference efficiency by shifting structural search to pretraining, but their one-pass autoregressive decoding

EditSSC: Toward Editable Semantic Occupancy Scenes with Unconditional Diffusion Models

AgentsDGX agent

arXiv:2606.09273v1 Announce Type: new Abstract: 3D semantic scene generation is crucial for autonomous driving applications, yet most methods rely on complex 3D-specific architectures such as triplane

Efficient Minimal Solvers for Relative Pose Estimation in Autonomous Driving Applications

Model ReleasesDGX agent

arXiv:2606.09569v1 Announce Type: cross Abstract: With the advancement of visual sensing systems, computer vision is playing an increasingly important role in autonomous driving and robot navigation.

Efficient Minimal Solvers for Visual-Inertial Relative Pose Estimation in Multi-Camera Systems

Model ReleasesDGX agent

arXiv:2606.09477v1 Announce Type: new Abstract: Estimating the relative poses of multi-camera systems is a fundamental problem in computer vision, with critical applications in autonomous vehicles, mo

Efficient Onboard Vision-Language Inference in UAV-Enabled Low-Altitude Economy Networks via LLM-Enhanced Optimization

ResearchDGX agent

arXiv:2510.10028v2 Announce Type: replace-cross Abstract: The rapid advancement of Low-Altitude Economy Networks (LAENets) has enabled a variety of applications, including aerial surveillance, environ

Efficient Scaling of LLM Training with Flexible Context Parallelism

ApplicationsDGX agent

arXiv:2602.21788v2 Announce Type: replace-cross Abstract: Scaling long-context capabilities is crucial for Large Language Models (LLMs). However, real-world data contain a large number of sequences wi

Efficient Skill Grounding via Code Refactoring with Small Language Models

Local AiDGX agent

arXiv:2606.07999v1 Announce Type: new Abstract: Effective skill grounding is essential for deploying reusable skills in embodied agents, as even minor embodiment or environmental differences can rende

Efficient Traffic Prediction at Scale: A Systematic Study of STGCN Architectural Depth

ResearchDGX agent

arXiv:2606.09539v1 Announce Type: new Abstract: Spatio-temporal graph neural networks (STGNNs) have become the dominant approach for traffic prediction, yet their computational requirements pose chall

Ego-Pi: VLA Fine-Tuning for Ego-Centric Human and Robot Data

TutorialsDGX agent

arXiv:2606.08107v1 Announce Type: cross Abstract: Robotics faces a fundamental challenge of data scarcity. Unlike language or vision research, there is no internet-scale dataset for robotic manipulati

EgoAERO: Learning Dexterous Manipulation from a Single Egocentric Video without Object Assets

SafetyDGX agent

arXiv:2606.08057v1 Announce Type: cross Abstract: Egocentric RGB-D videos offer a natural source of human dexterous manipulation demonstrations, but existing data is difficult to use for robot learnin

EgoPriMo: Egocentric Motion Generation for Interactive Humanoid Control

ResearchDGX agent

arXiv:2606.08495v1 Announce Type: cross Abstract: Humanoid robots require whole-body motions that adapt to scene context, task requirements, and user intent. Motion tracking reproduces specified traje

EgoTactile: Learning Grasp Pressure for Everyday Objects from Egocentric Video

Model ReleasesDGX agent

arXiv:2606.09243v1 Announce Type: cross Abstract: Estimating full-hand grasp pressure from egocentric video is critical for immersive VR and robotic manipulation, yet dense tactile sensing often relie

EinSort: Sorting is All We Need for Tensorizing LLM

ResearchDGX agent

arXiv:2606.08565v1 Announce Type: cross Abstract: Tensor networks provide efficient representations for compressing large neural networks. By carefully designing shapes and topologies, they can signif

Embedded Graph Convolutional Networks for Real-Time Event Data Processing on SoC FPGAs

ResearchDGX agent

arXiv:2406.07318v3 Announce Type: replace Abstract: The utilisation of event cameras represents an important and swiftly evolving trend aimed at addressing the constraints of traditional video systems

Emergence of Context Characteristics Sensitivity in Large Language Models

TutorialsDGX agent

arXiv:2606.09525v1 Announce Type: cross Abstract: During instruction fine-tuning (IFT), large language models (LLMs) learn to follow instructions by using the provided context to answer a query. While

Emergence via Phase Transitions: Mechanism Landscapes and Universal Convergence Across Complex Systems

ResearchDGX agent

arXiv:2606.07563v1 Announce Type: cross Abstract: Across machine learning, biology, and physics, independently evolving systems often converge toward strikingly similar high-level structures despite r

Emergence World: A Platform for Evaluating Long-Horizon Multi-Agent Autonomy

Model ReleasesDGX agent

arXiv:2606.08367v1 Announce Type: cross Abstract: Most evaluations of LLM agents look like exams: a discrete task, a clean environment, a score in minutes or hours. We argue that this approach is mism

Emergent alignment and the projectability of ethical personas

SafetyDGX agent

arXiv:2606.09475v1 Announce Type: new Abstract: Work on `emergent misalignment' shows that finetuning LLMs on narrow tasks can induce broadly misaligned behavior. This supports the `persona selection'

Empowering Feed-Forward Reconstruction Models with Metric Scale via Satellite Images

Local AiDGX agent

arXiv:2606.08205v1 Announce Type: new Abstract: Feed-forward 3D reconstruction models have recently shown strong generalization across diverse scenes, yet most of them recover geometry only up to an u

Enabling KV Caching of Shared Prefix for Diffusion Language Models

ResearchDGX agent

arXiv:2606.07571v1 Announce Type: cross Abstract: Key-value (KV) caching for shared prefixes is essential for high-throughput large language model (LLM) serving, but it faces critical challenges in em

End-to-End Context Compression at Scale

Model ReleasesDGX agent

arXiv:2606.09659v1 Announce Type: cross Abstract: Long-context language model inference is bottlenecked by memory, as the KV cache grows with context length. Recent techniques to compress the KV cache

End-to-End Control of a Powered Knee-Ankle Prosthesis Towards Unified, Tuning-Free Assistance

ResearchDGX agent

arXiv:2606.07902v1 Announce Type: new Abstract: Powered prostheses conventionally rely on impedance controllers that require extensive manual tuning and explicit mode classification. In this work, we

End-to-End Optimization of Incoherent Imaging for Classification Under Detector-Limited Readout

ResearchDGX agent

arXiv:2606.09792v1 Announce Type: new Abstract: End-to-end co-optimization of optical front-ends (e.g. metasurfaces) and neural network back-ends has been widely applied to imaging tasks, yet a formal

End-to-End Training for Discrete Token LLM based TTS System

Model ReleasesDGX agent

arXiv:2606.09234v1 Announce Type: cross Abstract: Recent state-of-the-art (SOTA) text-to-speech (TTS) systems typically adopt a cascaded pipeline consisting of a speech tokenizer, an autoregressive la

Engagement Process: Rethinking the Temporal Interface of Action and Observation

AgentsDGX agent

arXiv:2605.11484v2 Announce Type: replace Abstract: Task completion in digital and physical environments increasingly involves complex temporal interaction, where actions and observations unfold over

Enhanced Detection of Tiny Objects in Aerial Images

Model ReleasesDGX agent

arXiv:2509.17078v3 Announce Type: replace Abstract: While one-stage detectors like YOLOv8 offer fast training speed, they often under-perform on detecting small objects as a trade-off. This becomes ev

Enhancing Adversarial Robustness with Signed Distance Fields for Harmonizing Geometric Invariance and Texture

Local AiDGX agent

arXiv:2602.05175v2 Announce Type: replace Abstract: Deep neural networks demonstrate impressive performance in visual recognition but remain highly vulnerable to imperceptible adversarial attacks. Exi

Enhancing AI Interpretability and Safety through Localised Architectures

SafetyDGX agent

arXiv:2606.07998v1 Announce Type: cross Abstract: Recent advances in generative AI, especially powerful Large Language Models (LLMs) and Large Reasoning Models (LRMs), raise concerns over the interpre

← Previous
1…639640641642643…1517
Next →