AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “applications”

GridTimelineEvolution
13,198 results
19 May 2026

AASIST3: KAN-Enhanced AASIST Speech Deepfake Detection using SSL Features and Additional Regularization for the ASVspoof 2024 Challenge

ResearchDGX agent

arXiv:2408.17352v2 Announce Type: replace-cross Abstract: Automatic Speaker Verification (ASV) systems, which identify speakers based on their voice characteristics, have numerous applications, such a

ATRACT: A Trustworthy Robotic Autonomous system to support Casualty Triage

AgentsDGX agent

arXiv:2605.17123v1 Announce Type: cross Abstract: At a time when drones are increasingly associated with hostile operations, we re-purpose them for humanitarian and life-saving applications. However,

Augmenting Human Evaluation with LLM Judges: How Many Human Reviews Do You Need?

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.16354v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as automated evaluators of AI systems, including in high-stakes applications. In this role, LLMs ar

Avoiding Structural Failure Modes in Tabular Fair SSL: Online Primal-Dual Allocation under Confidence Gating

SafetyDGX agent

arXiv:2605.16446v1 Announce Type: cross Abstract: Semi-supervised learning (SSL) enables prediction with limited labels, but high-stakes tabular applications (medical, credit, recidivism) require stat

Axial-Relation Guided Fusion State Space Model for Optical-Elevation Sensing Image Segmentation

Local AiDGX agent

arXiv:2605.16768v1 Announce Type: new Abstract: Semantic segmentation of multi-source remote sensing images is a fundamental task for Earth observation applications. Existing methods often struggle wi

Best-of-Both-Worlds Multi-Dueling Bandits: Unified Algorithms for Stochastic and Adversarial Preferences under Condorcet and Borda Objectives

ResearchDGX agent

arXiv:2603.18972v3 Announce Type: replace Abstract: Multi-dueling bandits, where a learner selects m geq 2 arms per round and observes only the winner, arise naturally in many applications including r

Beyond Policy Optimization: A Data Curation Flywheel for Sparse-Reward Long-Horizon Planning

SafetyDGX agent

arXiv:2508.03018v2 Announce Type: replace Abstract: Large Language Reasoning Models have demonstrated remarkable success on static tasks, yet their application to multi-round agentic planning in inter

Bundle Adjustment in the Eager Mode

HardwareDGX agent

arXiv:2409.12190v4 Announce Type: replace-cross Abstract: Bundle adjustment (BA) is a critical technique in various robotic applications such as simultaneous localization and mapping (SLAM), augmented

CATA: Continual Machine Unlearning via Conflict-Averse Task Arithmetic

ResearchDGX agent

arXiv:2605.18610v1 Announce Type: cross Abstract: Vision-language models (VLMs) have shown remarkable ability in aligning visual and textual representations, enabling a wide range of multimodal applic

CATRF: Codec-Adaptive TriPlane Radiance Fields for Volumetric Content Delivery

ResearchDGX agent

arXiv:2605.18054v1 Announce Type: cross Abstract: Volumetric media promises next-generation content delivery applications, but its bandwidth demand remains a key bottleneck. Implicit and hybrid volume

Code-as-Room: Generating 3D Rooms from Top-Down View Images via Agentic Code Synthesis

Model ReleasesDGX agent

arXiv:2605.18451v1 Announce Type: new Abstract: Designing realistic and functional 3D indoor rooms is essential for a wide range of applications, including interior design, virtual reality, gaming, an

Context Memorization for Efficient Long Context Generation

Model ReleasesDGX agent

arXiv:2605.18226v1 Announce Type: cross Abstract: Modern large language model (LLM) applications increasingly rely on long conditioning prefixes to control model behavior at inference time. While pref

Counting Machine Parts

ResearchDGX agent

arXiv:2605.17952v1 Announce Type: new Abstract: Counting objects in an image is a task applicable across many domains. For instance, crowd counting, inventory counting, and cell counting have been the

DeMa: Dual-Path Delay-Aware Mamba for Efficient Multivariate Time Series Analysis

ResearchDGX agent

arXiv:2601.05527v2 Announce Type: replace-cross Abstract: Accurate and efficient multivariate time series (MTS) analysis is increasingly critical for a wide range of intelligent applications. Within t

Density-Ratio Weighted Behavioral Cloning: Learning Control Policies from Corrupted Datasets

SafetyDGX agent

arXiv:2510.01479v2 Announce Type: replace Abstract: Offline reinforcement learning (RL) enables policy optimization from fixed datasets, making it suitable for safety-critical applications where onlin

DeTrack: A Benchmark and Altitude-Aware Dual World Model for Drone-embodied Tracking

Model ReleasesDGX agent

arXiv:2605.17451v1 Announce Type: new Abstract: Aerial object tracking has broad applications in public safety, emergency rescue, wildlife monitoring, and related fields. However, existing aerial trac

Everything Google Cloud customers need to know coming out of Google I/O

Model ReleasesDGX agent

At Google Cloud Next ‘26, we unveiled the blueprint for the Agentic Enterprise, sharing our eighth-generation TPUs, Gemini Enterprise Agent Platform, a fully reimagined Agentic Data Cloud, Workspace I

FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech

Model ReleasesDGX agent

arXiv:2503.16492v3 Announce Type: replace-cross Abstract: ffective Human-Robot Interaction (HRI) is crucial for enhancing accessibility and usability in real-world robotics applications. However, exis

Flow-Direct: Feedback-Efficient and Reusable Guidance for Flow Models via Non-Parametric Guidance Field

ResearchDGX agent

arXiv:2605.16348v1 Announce Type: cross Abstract: Training-free guidance enables pre-trained diffusion and flow models to optimize application-specific objectives using feedback from external black-bo

FOL2NS: Generating Natural Sentences from First-Order Logic

ResearchDGX agent

arXiv:2605.18155v1 Announce Type: new Abstract: Translating formal language into natural language is a foundational challenge in NLP, driving various downstream applications in semantic parsing, theor

From BERT to T5: A Study of Named Entity Recognition

ResearchDGX agent

arXiv:2605.18462v1 Announce Type: new Abstract: Named entity recognition (NER) has been one of the essential preliminary steps in modern NLP applications. This report focuses on implementing the NER t

From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models

Model ReleasesDGX agent

arXiv:2508.01608v2 Announce Type: replace Abstract: Image geolocalization, the task of identifying the geographic location depicted in an image, is important for applications in crisis response, digit

Haptic Rendering of Fractional-Order Viscoelasticity: Passivity and Rendering Fidelity

ResearchDGX agent

arXiv:2605.16389v1 Announce Type: cross Abstract: Haptic rendering of viscoelastic materials that exhibit creep and stress relaxation is crucial for many applications, such as medical training with re

Herding CATs: ALARA for Agent Harness Engineering in Portable Composable Multi-Agent Teams

Local AiDGX agent

arXiv:2603.20380v2 Announce Type: replace-cross Abstract: Industry practitioners and academic researchers regularly use multi-agent systems to accelerate their work, but the applications through which

Here’s why Elon Musk lost his suit against OpenAI

ResearchDGX agent

On Monday, the jury in Musk v. Altman dealt Elon Musk a major blow—reaching a unanimous advisory verdict that he had sued OpenAI too late and, as a result, his claims are barred by the applicable stat

Heterogeneous Tasks Offloading in Vehicular Edge Computing: A Federated Meta Deep Reinforcement Learning Approach

SafetyDGX agent

arXiv:2605.18437v1 Announce Type: new Abstract: Vehicular edge computing (VEC) enables latency-sensitive vehicular applications by offloading computation-intensive tasks to nearby edge servers. Howeve

HierEdit: Region-Aware Hierarchical Diffusion for Efficient High-Resolution Editing

Local AiDGX agent

arXiv:2605.17294v1 Announce Type: new Abstract: High-resolution image editing is essential for professional and creative applications, yet existing multimodal diffusion-based editors remain computatio

How Do Electrocardiogram Models Scale?

Model ReleasesDGX agent

arXiv:2605.17276v1 Announce Type: cross Abstract: While scaling laws have established a fundamental framework for foundation models in natural language processing, their applicability to electrocardio

HyperTea: A Hypergraph-based Temporal Enhancement and Alignment Network for Moving Infrared Small Target Detection

Local AiDGX agent

arXiv:2508.10678v2 Announce Type: replace Abstract: In practical application scenarios, moving infrared small target detection (MIRSTD) remains highly challenging due to the target's small size, weak

LangChain Applied AI Engineer @palashshah takes us under the hood of LangSmith Engine.

AgentsDGX agent

LangSmith Engine is a tool within the LangChain ecosystem that provides visibility and debugging capabilities for AI applications built with LangChain. This technical deep-dive likely covers how the e

Learning Lifted Action Models from Traces with Minimal Information About Actions and States

ResearchDGX agent

arXiv:2605.18627v1 Announce Type: new Abstract: It has been recently shown that lifted STRIPS models can be learned correctly and efficiently from action traces alone; i.e., applicable action sequence

Lever: Speculative LLM Inference on Smartphones

ResearchDGX agent

arXiv:2605.16786v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly needed for interactive mobile applications, but high-quality models exceed the limited DRAM available on s

MetaLab: Few-Shot Game Changer for Image Recognition

ResearchDGX agent

arXiv:2507.22057v2 Announce Type: replace Abstract: Difficult few-shot image recognition has significant application prospects, yet remaining the substantial technical gaps with the conventional large

Mind the Gap: Learning Modality-Agnostic Representations with a Cross-Modality UNet

TutorialsDGX agent

arXiv:2605.16887v1 Announce Type: new Abstract: Cross-modality recognition has many important applications in science, law enforcement and entertainment. Popular methods to bridge the modality gap inc

MIRAGE: Robust multi-modal architectures translate fMRI-to-image models from vision to mental imagery

Model ReleasesDGX agent

arXiv:2605.17198v1 Announce Type: cross Abstract: To be useful for downstream applications, vision decoding models that are trained to reconstruct seen images from human brain activity must be able to

MoCA3D: Monocular 3D Bounding Box Prediction in the Image Plane

ResearchDGX agent

arXiv:2603.19538v2 Announce Type: replace Abstract: Monocular 3D object understanding has largely been cast as a 2D RoI-to-3D box lifting problem. However, emerging downstream applications require ima

MUSE: Multimodal Uncertainty Quantification of State Estimation

AgentsDGX agent

arXiv:2605.17421v1 Announce Type: new Abstract: Accurate visual state estimation has been a central topic in robotics with a wide range of applications in robot navigation, autonomous driving, and aut

New dances added to the reachy mini desktop app. Try them and post a video below?

IndustryDGX agent

New dance animations have been added to the Reachy Mini desktop application, a robotic arm control software. Users are invited to test the new dances and share video demonstrations of the features in

Offline Contextual Bandits in the Presence of New Actions

Model ReleasesDGX agent

arXiv:2605.18509v1 Announce Type: new Abstract: Automated decision-making algorithms drive applications such as recommendation systems and search engines. These algorithms often rely on off-policy con

OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents

Model ReleasesDGX agent

arXiv:2506.16042v2 Announce Type: replace Abstract: Generative AI is being leveraged to solve a variety of computer-use tasks involving desktop applications. State-of-the-art systems have focused sole

Pairwise Preference Reward and Group-Based Diversity Enhancement for Superior Open-Ended Generation

SafetyDGX agent

arXiv:2605.18191v1 Announce Type: new Abstract: Current reinforcement learning(RL) methods are broadly applicable and powerful in verifiable settings where scalar rewards can be provided. However, in

@Railway https://status.railway.com/

ResearchDGX agent

Railway is a platform for deploying and managing applications in the cloud, and this status page link provides real-time information about the service's operational status and any ongoing incidents or

Scalable Bi-causal Optimal Transport via KL Relaxation and Policy Gradients

SafetyDGX agent

arXiv:2605.17271v1 Announce Type: cross Abstract: Bi-causal optimal transport (OT) is a natural framework for comparing and coupling stochastic processes under nonanticipative information constraints,

Scientific Logicality Enriched Methodology for LLM Reasoning: A Practice in Physics

ResearchDGX agent

arXiv:2605.17104v1 Announce Type: new Abstract: With the continuous advancement of reasoning abilities in Large Language Models (LLMs), their application to scientific reasoning tasks has gained signi

Semantic Smoothing via Novel View Synthesis for Robust SAR Image Classification

SafetyDGX agent

arXiv:2605.16440v1 Announce Type: cross Abstract: Deep neural networks are vulnerable to adversarial perturbations, limiting deployment in safety-critical applications such as synthetic aperture radar

Some fun Gemini Omni use cases from the community👇🧵 (We’ll keep updating this thread throughout the day)

Model ReleasesDGX agent

This X thread from Google AI showcases practical and creative applications of Gemini Omni, Google's multimodal AI model, as demonstrated and shared by the user community. The thread appears to be a cu

StreamingEffect: Real-Time Human-Centric Video Effect Generation

HardwareDGX agent

arXiv:2605.17019v1 Announce Type: new Abstract: Streaming video effect generation is highly desirable for live human-centric applications such as e-commerce streaming, entertainment, and vlogging, yet

Structured Neural Marked Point Processes for Interpretable Event Interaction Modeling

Model ReleasesDGX agent

arXiv:2605.17568v1 Announce Type: new Abstract: Multi-class event streams arise in numerous real-world applications, where uncovering structured, interpretable inter-event relationships, together with

STT-Arena: A More Realistic Environment for Tool-Using with Spatio-Temporal Dynamics

Model ReleasesDGX agent

arXiv:2605.18548v1 Announce Type: cross Abstract: Large language models (LLMs) deployed in real-world agentic applications must be capable of replanning and adapting when mid-task disruptions invalida

Systematic Evaluation of the Quality of Synthetic Clinical Notes Rephrased by LLMs at Million-Note Scale

ResearchDGX agent

arXiv:2605.17775v1 Announce Type: cross Abstract: Large language models (LLMs) can generate or synthesize clinical text for a wide range of applications, from improving clinical documentation to augme

TClone: Low-Latency Forking of Live GUI Environments for Computer-Use Agents

SafetyDGX agent

arXiv:2605.17320v1 Announce Type: cross Abstract: Computer-use agents increasingly operate inside live personal workspaces, where their actions can modify files, applications, GUI state, credentials,

Voice ''Cloning'' is Style Transfer

ResearchDGX agent

arXiv:2605.16578v1 Announce Type: cross Abstract: Artificially generated speech is increasingly embedded in everyday life. Voice cloning in particular enables applications where identity preservation

what if we mapped older distributed systems patterns (like actor models or reactive, state-driven blackboards) to LLM agents?

AgentsDGX agent

This post explores conceptual parallels between classical distributed systems design patterns—such as actor models and reactive blackboard architectures—and their potential application to LLM agent de

You Had One Job: Per-Task Quantization Using LLMs' Hidden Representations

ResearchDGX agent

arXiv:2511.06516v3 Announce Type: replace Abstract: Many LLM applications require only narrow capabilities, yet standard post-training quantization (PTQ) methods allocate precision without considering

18 May 2026

Adapting Foundation Vision-Language Models to Medical Diagnosis via Query-Driven Expert Bridging

Model ReleasesDGX agent

arXiv:2505.21698v3 Announce Type: replace Abstract: Vision-language foundation models achieve promising performance in natural image classification, yet their direct application to medical imaging is

Beyond the Query: 5 Scenarios Laying the Foundation for the Agentic Era

Model ReleasesDGX agent

Accessing enterprise data is shifting from static reports to dynamic use by autonomous systems. To keep up, organizations must route fragmented data from SaaS, IoT, and legacy sources into secure, sca

CIS-BWE: Chaos-Informed Speech Bandwidth Extension

Model ReleasesDGX agent

arXiv:2507.15970v3 Announce Type: replace-cross Abstract: Recovering high-frequency components lost to bandwidth constraints is crucial for applications ranging from telecommunications to high-fidelit

Hardware-Software Co-Design of Scalable, Energy-Efficient Analog Recurrent Computations

Model ReleasesDGX agent

arXiv:2605.15216v1 Announce Type: cross Abstract: Always-on AI applications, from environmental sensors to biomedical implants, require ultra-low power consumption. Analog circuits offer a path to sub

it’s funny how people here just make stuff up.

SafetyDGX agent

it’s funny how people here just make stuff up. @GaryMarcus Take away LLM & every single AI application today goes back to the stone age, driverless cars will immediately break down, all the apps will

Learning Dynamic Structural Specialization for Underwater Salient Object Detection

Model ReleasesDGX agent

arXiv:2605.15535v1 Announce Type: new Abstract: Underwater salient object detection (USOD) has attracted increasing attention for underwater visual scene understanding and vision-guided robotic applic

← Previous
1…117118119120121…220
Next →