AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
64,475 results
Agents

Meridian: Metric-Semantic Primitive Matching for Cross-View Geo-Localization Beyond Urban Environments

DGX agent

arXiv:2606.06312v1 Announce Type: new Abstract: Successful robot automation requires accurate global localization to support repeatability, task planning, goal specification, and safe operation. Howev

agentsarxiv-cs-ro
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

MIRAI: Prediction and Generation of High-Impact Academic Research

DGX agent

arXiv:2606.05443v1 Announce Type: cross Abstract: The rapid pace of scientific publishing has made the identification and synthesis of high-impact work an increasingly urgent challenge. We introduce M

researcharxiv-cs-cl
5 Jun 2026
Agents

MLEvolve: A Self-Evolving Framework for Automated Machine Learning Algorithm Discovery

DGX agent

arXiv:2606.06473v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly applied to long-horizon tasks such as scientific discovery and machine learning engineering (MLE),

agentsarxiv-cs-cl
5 Jun 2026
Safety

MoDex: A Diffusion Policy for Sequential Multi-Object Dexterous Grasping

DGX agent

arXiv:2606.05407v1 Announce Type: new Abstract: This work addresses sequentially grasping multiple objects with a single dexterous hand without releasing those already held. Most dexterous grasping me

safetyarxiv-cs-ro
5 Jun 2026
Research

Monte Carlo Steklov Operators for Large-Scale Geometry Processing in the Wild

DGX agent

arXiv:2606.05581v1 Announce Type: cross Abstract: Intrinsic methods fill the default toolbox for geometry processing on meshes. Intrinsic operators, in particular the Laplacian, underlie methods that

researcharxiv-cs-cv
5 Jun 2026
Applications

MotionDisco: Motion Discovery for Extreme Humanoid Loco-Manipulation

DGX agent

arXiv:2606.06139v1 Announce Type: new Abstract: We present MotionDisco, a framework that discovers contact-rich, long-horizon humanoid loco-manipulation motions from scratch, without relying on teleop

applicationsarxiv-cs-ro
5 Jun 2026
Research

MPCoT: Reward-Guided Multi-Path Latent Reasoning for Test-Time Scalable Vision-Language-Action

DGX agent

arXiv:2606.06245v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies remain brittle in long-horizon and high-uncertainty control, where one-pass action decoding provides limited infer

researcharxiv-cs-ro
5 Jun 2026
Research

MS-DKC: A Dataset Knowledge Card Framework for Designing and Adapting Medical Image Segmentation Models

DGX agent

arXiv:2606.06103v1 Announce Type: new Abstract: Medical image segmentation is often framed as a search for stronger architectures, but this can obscure a more fundamental question: what does the datas

researcharxiv-cs-cv
5 Jun 2026
Research

Multi-Granularity Reasoning for Natural Language Inference

DGX agent

arXiv:2606.05181v1 Announce Type: new Abstract: Natural Language Inference (NLI) is a fundamental task in natural language understanding that requires determining the logical relationship between a pr

researcharxiv-cs-cl
5 Jun 2026
Safety

Multi-Resolution Tactile Imitation Learning for Contact-Rich Robotic Manipulation

DGX agent

arXiv:2606.06281v1 Announce Type: new Abstract: Touch sensing is beneficial for solving a wide variety of manipulation tasks. While there exists a wide range of tactile sensors with different properti

safetyarxiv-cs-ro
5 Jun 2026
Research

Multi-Task Crack Foundation Model for Engineering-Reliable Crack Representation and Topology Preservation in Civil Infrastructure

DGX agent

arXiv:2606.05641v1 Announce Type: new Abstract: Reliable crack assessment requires not only accurate pixel-level masks but also connected crack geometry and confidence estimates that remain stable und

researcharxiv-cs-cv
5 Jun 2026
Research

Multi-task Learning is Not Enough: Representational Entanglement in Dual-output Second Language Speech Recognition

DGX agent

arXiv:2606.06065v1 Announce Type: new Abstract: Second-language (L2) speech recognition often requires transcriptions of pronunciations and intended meanings. Multi-task learning (MTL) is a natural ap

researcharxiv-cs-cl
5 Jun 2026
Research

Multilingual Coreference Resolution via Cycle-Consistent Machine Translation

DGX agent

arXiv:2606.05444v1 Announce Type: new Abstract: Coreference resolution is a core NLP task, having a broad range of downstream applications, e.g.~machine translation, question answering, document summa

researcharxiv-cs-cl
5 Jun 2026
Research

Multilingual Detection of Alzheimer's Disease from Speech: A Cross-Linguistic Transfer Learning Approach

DGX agent

arXiv:2606.05545v1 Announce Type: new Abstract: The development of multilingual Alzheimer's Disease Dementia (AD) detection models presents significant challenges due to the resource-intensive and tim

researcharxiv-cs-cl
5 Jun 2026
Research

Multimodal Sexism Identification and Characterization using Large Language Models and Gradient Boosting

DGX agent

arXiv:2606.05997v1 Announce Type: new Abstract: We present the AILS-NTUA submission to the EXIST 2026 Lab at CLEF, addressing multimodal sexism identification and characterization in memes (Task 2) an

researcharxiv-cs-cv
5 Jun 2026
Research

Narrative Knowledge Weaver: Narrative-Centric Retrieval-Augmented Reasoning for Long-Form Text Understanding

DGX agent

arXiv:2606.05724v1 Announce Type: new Abstract: Long-form narrative QA requires reasoning over evolving story worlds rather than isolated passages: answers may depend on earlier goals, changing charac

researcharxiv-cs-cl
5 Jun 2026
Local Ai

NAVIRA: Decoupled Stochastic Remasking for Masked Diffusion Language Models

DGX agent

arXiv:2606.06031v1 Announce Type: new Abstract: Masked diffusion language models generate text by iteratively unmasking many tokens in parallel, but this speed comes with a correction problem: tokens

local-aiarxiv-cs-cl
5 Jun 2026
Research

Next-Generation Parallel Decoder for LPDR: Architectural Optimization and Class-Balanced GAN-Augmentation

DGX agent

arXiv:2606.05785v1 Announce Type: new Abstract: Real-Time License Plate Detection and Recognition (LPDR) forms the backbone of modern smart cities. Although the YOLOV5-PDLPR model substantially improv

researcharxiv-cs-cv
5 Jun 2026
Research

NIV: Neural Axis Variations for Variable Font Generation

DGX agent

arXiv:2606.05261v1 Announce Type: new Abstract: Variable fonts enable continuous variation of glyph geometry along semantic design axes such as weight, width, slant, and optical size. However, constru

researcharxiv-cs-cv
5 Jun 2026
Research

Noise-Adaptive Regularization for Robust Multi-Label Remote Sensing Image Classification

DGX agent

arXiv:2601.08446v2 Announce Type: replace Abstract: The development of reliable methods for multi-label classification (MLC) has become a prominent research direction in remote sensing (RS). As the sc

researcharxiv-cs-cv
5 Jun 2026
Model Releases

Noise-Aware Visual Representation Learning for Medical Visual Question Answering

DGX agent

arXiv:2606.05535v1 Announce Type: new Abstract: Medical visual question answering (Med-VQA) has strong potential for clinical decision support by enabling AI models to interpret medical images and ans

model-releasesarxiv-cs-cv
5 Jun 2026
Agents

OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions

DGX agent

arXiv:2602.05843v2 Announce Type: replace Abstract: The rapid advancement of Large Language Models (LLMs) has catalyzed the development of autonomous agents capable of navigating complex environments.

agentsarxiv-cs-cl
5 Jun 2026
Model Releases

Oklch+: A Three-Parameter Extension of Oklab for Improved Color Difference Prediction

DGX agent

arXiv:2606.05255v1 Announce Type: cross Abstract: Oklab and its cylindrical representation Oklch are widely adopted in interpolation and design workflows as perceptually motivated color spaces, but th

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

OLIVE: Online Low-Rank Incremental Learning for Efficient Adaptive Exoskeletons

DGX agent

arXiv:2606.05234v1 Announce Type: new Abstract: Wearable exoskeleton systems hold promise for restoring mobility in individuals with physical impairments, yet most existing controllers rely on static

model-releasesarxiv-cs-ro
5 Jun 2026
Safety

On Advantage Estimates for Max@K Policy Gradients

DGX agent

arXiv:2606.06080v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards is widely used for post-training reasoning models, but sparse outcome rewards make exploration difficul

safetyarxiv-cs-cl
5 Jun 2026
Applications

OneReason Technical Report

DGX agent

arXiv:2606.06260v1 Announce Type: cross Abstract: Generative recommendation models in the OneRec family have been widely deployed in many real-world services, such as short-video, live-streaming, adve

applicationsarxiv-cs-cl
5 Jun 2026
Model Releases

Operation-Guided Progressive Human-to-AI Text Transformation Benchmark for Multi-Granularity AI-Text Detection

DGX agent

arXiv:2606.06481v1 Announce Type: new Abstract: As AI writing assistants become increasingly integrated into real-world drafting and revision workflows, many documents are no longer purely human-writt

model-releasesarxiv-cs-cl
5 Jun 2026
Research

ORACLE-CT: Anatomy-Aware Support Pooling for CT Classification

DGX agent

arXiv:2606.05460v1 Announce Type: new Abstract: Abdominal CT disease classification is challenging because each scan is a large 3D volume with many possible findings, while diagnostic evidence is ofte

researcharxiv-cs-cv
5 Jun 2026
Safety

OrderGrad: Optimizing Beyond the Mean with Order-Statistic Policy Gradient Estimation

DGX agent

arXiv:2606.06096v1 Announce Type: cross Abstract: Policy-gradient methods usually optimize expected return, but many real world applications care about distributional properties of returns: tail risk,

safetyarxiv-cs-cl
5 Jun 2026
Safety

Ousiometrics: The essence of meaning aligns with a power-danger-structure framework instead of valence-arousal-dominance

DGX agent

arXiv:2110.06847v3 Announce Type: replace Abstract: From work emerging through the middle of the 20th century, the essence of meaning has become widely accepted as being described by the three orthogo

safetyarxiv-cs-cl
5 Jun 2026
Applications

Ouvia: A User-centered Framework for Measuring Usability of Speech Translation in Real-World Communication Scenarios

DGX agent

arXiv:2606.06177v1 Announce Type: new Abstract: Speech translation (ST) is increasingly adopted in user applications, yet its evaluation largely focuses on decontextualized testbeds and holistic quali

applicationsarxiv-cs-cl
5 Jun 2026
Research

PAR3D: A Unified 3D-MLLM with Part-Aware Representation for Scene Understanding

DGX agent

arXiv:2606.06485v1 Announce Type: new Abstract: Recent advances in 3D multimodal large language models (3D-MLLMs) have enabled unified solutions for 3D scene understanding tasks, including visual ques

researcharxiv-cs-cv
5 Jun 2026
Research

Parallel Jacobi Decoding for Fast Autoregressive Image Generation

DGX agent

arXiv:2606.05703v1 Announce Type: new Abstract: Autoregressive (AR) models have demonstrated remarkable performance in generating high-fidelity images. However, their inherently sequential next-token

researcharxiv-cs-cv
5 Jun 2026
Agents

PathWISE: Multi-Agent Cancer Pathway Triaging Ontology Learning from Clinical Flowcharts

DGX agent

arXiv:2605.25970v2 Announce Type: replace Abstract: Clinical pathways are disseminated as visual flowcharts where spatial topology, arrow direction, colour coding, and font weight encode critical tria

agentsarxiv-cs-cv
5 Jun 2026
Safety

PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation

DGX agent

arXiv:2503.14295v3 Announce Type: replace Abstract: Recent advancements in audio-driven talking face generation have made great progress in lip synchronization. However, current methods often lack suf

safetyarxiv-cs-cv
5 Jun 2026
Model Releases

PEFT of SLM for Telecommunications Customer Support: A Comparative Study of LoRA Configurations with Energy Consumption Analysis

DGX agent

arXiv:2606.05176v1 Announce Type: new Abstract: While large language models (LLMs) show strong performance in natural language understanding and generation, their evaluation and adaptation to domain-s

model-releasesarxiv-cs-cl
5 Jun 2026
Agents

Personal AI Agent for Camera Roll VQA

DGX agent

arXiv:2606.05275v1 Announce Type: new Abstract: We study the personal camera roll visual question answering setting. In this setting, a conversational AI assistant can access a user's personal camera

agentsarxiv-cs-cv
5 Jun 2026
Research

PHUMA: Physically Reliable Humanoid Locomotion Dataset

DGX agent

arXiv:2510.26236v2 Announce Type: replace Abstract: Motion imitation is a promising approach for humanoid locomotion, enabling agents to acquire humanlike behaviors. Existing methods typically rely on

researcharxiv-cs-ro
5 Jun 2026
Model Releases

Physics-Guided Deep Unfolding for Blind Cross-Sensor Spectral Super-Resolution via Learning the Spectral Transformation Function

DGX agent

arXiv:2606.05759v1 Announce Type: new Abstract: Hyperspectral imaging provides rich spectral information for quantitative remote sensing, yet hyperspectral sensors remain costly and thus unavailable i

model-releasesarxiv-cs-cv
5 Jun 2026
Research

Physics in 2-Steps: Locking Motion Priors Before Visual Refinement Erases Them

DGX agent

arXiv:2606.06361v1 Announce Type: new Abstract: Image-to-Video diffusion models leverage input images to generate visually stunning content, yet frequently produce motion that violates physical laws.

researcharxiv-cs-cv
5 Jun 2026
Safety

PiL-World: A Chunk-Wise World Model for VLA Policy-in-the-Loop Evaluation

DGX agent

arXiv:2606.05773v1 Announce Type: new Abstract: Vision-language-action (VLA) policies operate in a closed loop in real-world robot tasks: a robot observes the scene, executes an action chunk, and cond

safetyarxiv-cs-ro
5 Jun 2026
Safety

Pitfalls of Evaluating Language Models with Open Benchmarks

DGX agent

arXiv:2507.00460v3 Announce Type: replace Abstract: Open Large Language Model (LLM) benchmarks, such as HELM and BIG-Bench, provide standardized and transparent evaluation protocols that support compa

safetyarxiv-cs-cl
5 Jun 2026
Agents

PLAN-S: Bridging Planning with Latent Style Dynamics for Autonomous Driving World Models

DGX agent

arXiv:2606.06014v1 Announce Type: cross Abstract: Latent world models (LWMs) have strengthened end-to-end autonomous driving by forecasting compact scene dynamics for downstream planning. However, exi

agentsarxiv-cs-ro
5 Jun 2026
Model Releases

PlanBench-V: A Spatial Planning Map Benchmark for Vision-Language Models

DGX agent

arXiv:2606.05744v1 Announce Type: new Abstract: Spatial planning maps are central to territorial governance, translating planning objectives, regulations, and spatial strategies into visual forms for

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

Predict and Reconstruct: Joint Objectives for Self-Supervised Language Representation Learning

DGX agent

arXiv:2606.05173v1 Announce Type: new Abstract: Masked language modelling (MLM) has been the dominant pre-training objective for text encoders since BERT, yet it encourages representations that are st

model-releasesarxiv-cs-cl
5 Jun 2026
Research

Predictable Scaling Laws of Optimal Hyperparameters for LLM Continued Pre-training

DGX agent

arXiv:2606.05610v1 Announce Type: new Abstract: The efficacy of continued pre-training for Large Language Models (LLMs) hinges upon hyperparameter configurations, such as learning rate and batch size.

researcharxiv-cs-cl
5 Jun 2026
Research

Preserving Full 6-DOF Actuation Under Abrupt Total Rotor Failures: Passive Fault-Tolerant Flight Control Using a Biaxial-Tilt Hexacopter

DGX agent

arXiv:2606.05663v1 Announce Type: new Abstract: Conventional multirotors suffer from a rapid collapse of attainable wrench space (AWS) under abrupt total rotor failures, rendering full 6-DOF recovery

researcharxiv-cs-ro
5 Jun 2026
Local Ai

ProSarc: Prosody-Aware Sarcasm Recognition Framework via Temporal Prosodic Incongruity

DGX agent

arXiv:2606.06168v1 Announce Type: cross Abstract: We present ProSarc, an audio-only framework that detects sarcasm by modelling temporal prosodic incongruity, that is, the mismatch between local proso

local-aiarxiv-cs-cl
5 Jun 2026
← Previous
1…627628629630631…1344
Next →