AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
64,475 results
Model Releases

FATE: Focal-modulated Attention Encoder for Multivariate Time-series Forecasting

DGX agent

arXiv:2408.11336v3 Announce Type: replace-cross Abstract: Climate change stands as one of the most pressing global challenges of the twenty-first century, with far-reaching consequences such as rising

model-releasesarxiv-cs-cv
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

FiLM-Based Speaker Conditioning of a SpeechLLM for Pathological Speech Recognition

DGX agent

arXiv:2606.06211v1 Announce Type: new Abstract: Automatic speech recognition (ASR) has advanced remarkably for standard speech; however, pathological speech from neurological conditions remains a sign

model-releasesarxiv-cs-cl
5 Jun 2026
Hardware

Flash-WAM: Modality-Aware Distillation for World Action Models

DGX agent

arXiv:2606.05254v1 Announce Type: cross Abstract: World-action models (WAMs) jointly generate future video and robot actions through iterative diffusion, achieving strong performance on manipulation b

hardwarearxiv-cs-cv
5 Jun 2026
Safety

Flow-based Policy Adaptation without Policy Updates

DGX agent

arXiv:2606.06461v1 Announce Type: new Abstract: Leveraging prior knowledge from pretrained policies, foundation models, or human operators offers an efficient alternative to learning robot skills from

safetyarxiv-cs-ro
5 Jun 2026
Safety

FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization

DGX agent

arXiv:2606.05468v1 Announce Type: new Abstract: Post-training Vision-Language-Action (VLA) models into policies that can be reliably deployed on real robots remains a major bottleneck. SFT and DAgger

safetyarxiv-cs-ro
5 Jun 2026
Research

FontFusion: Enhancing Generative Text in Diffusion Models with Typographic Conditioning

DGX agent

arXiv:2606.06066v1 Announce Type: new Abstract: Typography generation in diffusion models faces a persistent trade-off: enabling precise font control typically degrades text legibility, while maintain

researcharxiv-cs-cv
5 Jun 2026
Safety

Forgive or forget: Understanding the context of hate in audio retrieval systems

DGX agent

arXiv:2606.05857v1 Announce Type: new Abstract: Handling toxic retrieval in text-to-audio systems is challenging due to contextual dependencies. Existing strategies (e.g., rephrasing, summarization) r

safetyarxiv-cs-cl
5 Jun 2026
Tutorials

Formal Concept Lattices are Good Semantic Scaffolds for Concept-Based Learning

DGX agent

arXiv:2606.05471v1 Announce Type: new Abstract: Learning semantics is essential for deep learning models to be interpretable and better aligned with human reasoning. Concept-based models approach this

tutorialsarxiv-cs-cv
5 Jun 2026
Research

FOXGLOVE: Understanding Goal-Oriented and Anchored Writing Feedback from Experts and LLMs on Argumentative Essays

DGX agent

arXiv:2606.06271v1 Announce Type: new Abstract: While large language models (LLMs) are increasingly used to generate writing feedback, there remains no systematic comparison of LLM and expert feedback

researcharxiv-cs-cl
5 Jun 2026
Research

Framing, Judging, Steering: An Assessable Competency Model for Teach-ing Students to Reason With Generative AI

DGX agent

arXiv:2606.05983v1 Announce Type: cross Abstract: Generative AI makes answers easy and understanding hard, and uncritical use invites cognitive offloading. Schools still measure unaided performance, y

researcharxiv-cs-cl
5 Jun 2026
Research

From Scoring to Explanations: Evaluating SHAP and LLM Rationales for Rubric-based Teaching Quality Assessment

DGX agent

arXiv:2606.05180v1 Announce Type: new Abstract: Automated scoring models are increasingly used to assign rubric-based quality ratings to complex language performances, including classroom transcripts,

researcharxiv-cs-cl
5 Jun 2026
Model Releases

From Self to Other: Evaluating Demographic Perspective-Taking in LLM Hate Speech Annotation

DGX agent

arXiv:2606.06266v1 Announce Type: new Abstract: Hate speech detection is inherently subjective: people from different demographic groups perceive the same content very differently. Collecting enough a

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

FUSAR-GPT : A Spatiotemporal Feature-Embedded and Two-Stage Decoupled Visual Language Model for SAR Imagery

DGX agent

arXiv:2602.19190v4 Announce Type: replace Abstract: Research on the intelligent interpretation of all-weather, all-time Synthetic Aperture Radar (SAR) is crucial for advancing remote sensing applicati

model-releasesarxiv-cs-cv
5 Jun 2026
Research

Gender Artifacts from Art History to Text-to-Image Generation

DGX agent

arXiv:2606.05829v1 Announce Type: new Abstract: Artistic styles are rooted in specific socio-historical contexts that encode social hierarchies, including distinct constructions of gender. Yet in AI r

researcharxiv-cs-cv
5 Jun 2026
Model Releases

Generic Triple-Latent Compression with Gated Associative Retrieval

DGX agent

arXiv:2606.05175v1 Announce Type: new Abstract: We study generic triple-latent sequence models that maintain a running token state and compressed pair-memory pathway to capture higher-order token inte

model-releasesarxiv-cs-cl
5 Jun 2026
Local Ai

GenTract: Generative Global Tractography

DGX agent

arXiv:2511.13183v2 Announce Type: replace Abstract: Tractography is the process of inferring the trajectories of white-matter pathways in the brain from diffusion magnetic resonance imaging (dMRI). Lo

local-aiarxiv-cs-cv
5 Jun 2026
Tutorials

Geodesic Flow Matching on a Riemannian Degradation Manifold for Blind Image Restoration

DGX agent

arXiv:2606.06278v1 Announce Type: new Abstract: Blind image restoration requires recovering clean images from observations corrupted by unknown and potentially mixed degradations. While recent determi

tutorialsarxiv-cs-cv
5 Jun 2026
Safety

Geometry-Aware Dataset Condensation for Diffusion Model Training

DGX agent

arXiv:2606.05883v1 Announce Type: new Abstract: Dataset condensation aims to construct compact datasets from real data via synthesis or selection. However, existing approaches are ill-suited for diffu

safetyarxiv-cs-cv
5 Jun 2026
Safety

GLASS: GRPO-Trained LoRA for Acoustic Style Steering in Zero-Shot Text-to-Speech

DGX agent

arXiv:2606.05889v1 Announce Type: cross Abstract: We propose GLASS, a framework for composable acoustic style control in zero-shot autoregressive text-to-speech (TTS) that learns controls from post-ge

safetyarxiv-cs-cl
5 Jun 2026
Model Releases

Global Cross-Modal Geo-Localization: A Million-Scale Dataset and a Physical Consistency Learning Framework

DGX agent

arXiv:2603.08491v2 Announce Type: replace Abstract: Cross-modal Geo-localization (CMGL) matches ground-level text descriptions with geo-tagged aerial imagery, which is crucial for pedestrian navigatio

model-releasesarxiv-cs-cv
5 Jun 2026
Research

Global-Local Monte Carlo Tree Search in Vision-Language Models for Text-to-3D Indoor Scene Generation

DGX agent

arXiv:2606.06002v1 Announce Type: new Abstract: Large Vision-Language Models have achieved significant reasoning performance in various tasks.However, there are few studies on text-to-3D indoor scene

researcharxiv-cs-cv
5 Jun 2026
Research

GMBFormer: An NDVI-Guided Global Memory Bank Transformer for Urban Green-Space Extraction from Ultra-High-Resolution Imagery

DGX agent

arXiv:2606.06363v1 Announce Type: new Abstract: Urban green-space extraction from ultra-high-resolution (UHR) imagery is commonly performed patch by patch, which limits semantic reuse among spatially

researcharxiv-cs-cv
5 Jun 2026
Hardware

Gotta Grow Fast: Design and Benchmarking of a Tip Mount for High-Speed Vine Robots

DGX agent

arXiv:2606.06040v1 Announce Type: new Abstract: Soft, growing vine robots extend through tip eversion, a mechanism that enables navigation through cluttered environments. However, integrating cameras

hardwarearxiv-cs-ro
5 Jun 2026
Research

GRAMformer: Any-Order Modality Interactions via Volumetric Multimodal Cross-Attention

DGX agent

arXiv:2606.06249v1 Announce Type: new Abstract: Transformer-based multimodal models rely on attention mechanisms to integrate information across heterogeneous modalities. Despite their success, existi

researcharxiv-cs-cv
5 Jun 2026
Safety

Grounded but Misleading: Evaluating Semantic Alignment in AI-Generated Security Explanations

DGX agent

arXiv:2602.05056v2 Announce Type: replace-cross Abstract: Online scams increasingly leverage fluent and context-aware social engineering strategies, creating growing demand for AI systems that explain

safetyarxiv-cs-cl
5 Jun 2026
Hardware

GS-NFS: Bandwidth-adaptive Streaming of Dynamic Gaussian Splats and Point Clouds

DGX agent

arXiv:2606.05650v1 Announce Type: cross Abstract: Dynamic 3D Gaussian Splatting (3DGS) holds great promise as a 3D video streaming technology since it can represent complex 3D scenes with high fidelit

hardwarearxiv-cs-cv
5 Jun 2026
Safety

HANDOFF: Humanoid Agentic Task-Space Whole-Body Control via Distilled Complementary Teachers

DGX agent

arXiv:2606.06493v1 Announce Type: new Abstract: For a humanoid robot to be deployed in the real world, the choice of command space (i.e., the interface between task planning and whole-body control) is

safetyarxiv-cs-ro
5 Jun 2026
Model Releases

Harmonious Parameter Adaptation in Continual Visual Instruction Tuning for Safety-Aligned MLLMs

DGX agent

arXiv:2511.20158v2 Announce Type: replace Abstract: While continual visual instruction tuning (CVIT) has shown promise in adapting multimodal large language models (MLLMs), existing studies predominan

model-releasesarxiv-cs-cv
5 Jun 2026
Agents

Harnessing Generalist Agents for Contextualized Time Series

DGX agent

arXiv:2606.05404v1 Announce Type: cross Abstract: Time series are often embedded in rich contexts that are essential for holistic modeling. Moreover, real-world practitioners often require end-to-end

agentsarxiv-cs-cl
5 Jun 2026
Model Releases

Harnessing Structural Context for Entity Alignment Foundation Models

DGX agent

arXiv:2606.06109v1 Announce Type: new Abstract: Entity alignment (EA) aims to identify equivalent entities across heterogeneous knowledge graphs (KGs) and is a key component of knowledge fusion and cr

model-releasesarxiv-cs-cl
5 Jun 2026
Research

HDST-GNN: Heterogeneous Dynamic Spatiotemporal Graph Neural Networks for Multi-Object Tracking in UAV Aerial Imagery

DGX agent

arXiv:2606.05587v1 Announce Type: new Abstract: Multi-object tracking (MOT) from UAV imagery presents unique challenges: altitude varies across sequences, objects are small and densely packed, and fre

researcharxiv-cs-cv
5 Jun 2026
Safety

HERO: Learning Humanoid End-Effector Control for Visual Whole-Body Open-Vocabulary Object Grasping

DGX agent

arXiv:2602.16705v3 Announce Type: replace-cross Abstract: Visual loco-manipulation of arbitrary in-the-wild objects requires accurate end-effector (EE) control and a generalizable understanding of the

safetyarxiv-cs-cv
5 Jun 2026
Research

Hierarchical Mask-Enhanced Dual Reconstruction Network for Few-Shot Fine-Grained Image Classification

DGX agent

arXiv:2506.20263v2 Announce Type: replace Abstract: Few-shot fine-grained image classification (FS-FGIC) is challenging as it requires distinguishing visually similar subclasses with extremely limited

researcharxiv-cs-cv
5 Jun 2026
Model Releases

HOLO: Homography-Guided Pose Estimator Network for Fine-Grained Visual Localization on SD Maps

DGX agent

arXiv:2601.02730v3 Announce Type: replace Abstract: Visual localization on standard-definition (SD) maps has emerged as a promising low-cost and scalable solution for autonomous driving. However, exis

model-releasesarxiv-cs-cv
5 Jun 2026
Research

HomeWorld: A Unified Floorplan-to-Furnished Framework for Generating Controllable, Densely Interactive Whole-Home Scenes

DGX agent

arXiv:2606.06390v1 Announce Type: new Abstract: Indoor scene generation is crucial for robot simulation and modern interior design. However, complex layouts together with scarce 3D scene data make lea

researcharxiv-cs-cv
5 Jun 2026
Research

Horse Eye Blink Detection and Classification for Equine Affective State Assessment

DGX agent

arXiv:2606.05458v1 Announce Type: new Abstract: Automated detection of equine facial action units (AUs) is a promising yet under-explored avenue for pain and affective state assessment in horses. Half

researcharxiv-cs-cv
5 Jun 2026
Safety

Human Adults and LLMs as Scientists: Who Benefits from Active Exploration?

DGX agent

arXiv:2606.06464v1 Announce Type: new Abstract: A long-standing finding in the causal learning literature is that adults struggle to identify conjunctive causal rules, where an effect requires the sim

safetyarxiv-cs-cl
5 Jun 2026
Model Releases

Humans' ALMANAC: A Human Collaboration Dataset of Action-Level Mental Model Annotations for Agent Collaboration

DGX agent

arXiv:2606.06388v1 Announce Type: cross Abstract: Recent advances in LLM agents have enabled complex cognitive capabilities, such as multi-step reasoning, planning, and tool use, that increasingly pos

model-releasesarxiv-cs-cl
5 Jun 2026
Research

HyperVis: Continuous Latent Visual Relational Graphs on the Lorentz Hyperboloid for Compositional Reasoning

DGX agent

arXiv:2606.06100v1 Announce Type: new Abstract: Vision-Language Models (VLMs) struggle with compositional reasoning that requires understanding inter-object relationships. A natural remedy is to injec

researcharxiv-cs-cv
5 Jun 2026
Model Releases

IA-RAG: Interval-Algebra-Driven Temporal Reasoning for Dynamic Knowledge Retrieval

DGX agent

arXiv:2606.06044v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has shown strong effectiveness in grounding Large Language Models (LLMs) with external knowledge. However, existing

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

IDEAL: Leveraging Infinite and Dynamic Characterizations of Large Language Models for Query-focused Summarization

DGX agent

arXiv:2407.10486v3 Announce Type: replace-cross Abstract: Query-focused summarization (QFS) aims to produce summaries that answer particular questions of interest, enabling greater user control and pe

safetyarxiv-cs-cl
5 Jun 2026
Research

Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction

DGX agent

arXiv:2606.05769v1 Announce Type: new Abstract: Video event prediction (VEP) requires models to infer unobserved future states from partial video evidence. Existing video MLLMs usually verbalize inter

researcharxiv-cs-cv
5 Jun 2026
Model Releases

Improving Answer Extraction in Context-based Question Answering Systems Using LLMs

DGX agent

arXiv:2606.06197v1 Announce Type: new Abstract: Question answering (QA) systems have achieved notable progress with the advent of large language models (LLMs). However, they still face challenges in a

model-releasesarxiv-cs-cl
5 Jun 2026
Local Ai

Improving Heart-Focused Medical Question Answering in LLMs via Variance-Aware Rubric Rewards with GRPO

DGX agent

arXiv:2606.05174v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown strong promise in healthcare applications. Yet deploying general-purpose models in real-world settings remains d

local-aiarxiv-cs-cl
5 Jun 2026
Applications

In-Context Multiple Instance Learning

DGX agent

arXiv:2606.06458v1 Announce Type: cross Abstract: Multiple Instance Learning (MIL) addresses problems where supervision is available at the level of bags of instances and has been successfully applied

applicationsarxiv-cs-cv
5 Jun 2026
Research

InfoDensity: Rewarding Information-Dense Traces for Efficient Reasoning

DGX agent

arXiv:2603.17310v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) with extended reasoning capabilities often generate verbose and redundant reasoning traces, incurring unnecessary

researcharxiv-cs-cl
5 Jun 2026
Research

InfoShield: Privacy-Preserving Speech Representations for Mental Health Screening via Information-Theoretic Optimization

DGX agent

arXiv:2606.05561v1 Announce Type: new Abstract: Speech-based mental health screening offers scalable depression detection, yet clinical deployment faces a significant barrier: users' privacy concerns

researcharxiv-cs-cl
5 Jun 2026
Research

Interpreting Style Representations via Style-Eliciting Prompts

DGX agent

arXiv:2606.05716v1 Announce Type: new Abstract: Style representation learning is a powerful tool for authorship analysis and modeling writing style, yet the latent nature of learned representations ma

researcharxiv-cs-cl
5 Jun 2026
← Previous
1…625626627628629…1344
Next →