AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
Human
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
1 May 2026

GUI Agents with Reinforcement Learning: Toward Digital Inhabitants

AgentsDGX agent

arXiv:2604.27955v1 Announce Type: new Abstract: Graphical User Interface (GUI) agents have emerged as a promising paradigm for intelligent systems that perceive and interact with graphical interfaces

GuideDog: A Real-World Egocentric Multimodal Dataset for Blind and Low-Vision Accessibility-Aware Guidance

Model ReleasesDGX agent

arXiv:2503.12844v2 Announce Type: replace Abstract: For people affected by blindness and low vision (BLV), safe and independent navigation remains a major challenge, impacting over 2.2 billion individ

GVCC: Zero-Shot Video Compression via Codebook-Driven Stochastic Rectified Flow

ResearchDGX agent

arXiv:2603.26571v3 Announce Type: replace-cross Abstract: At ultra-low bitrates, high-fidelity reconstruction requires sampling plausible videos from the posterior rather than regressing to oversmooth

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

HATS: An Open data set Integrating Human Perception Applied to the Evaluation of Automatic Speech Recognition Metrics

ResearchDGX agent

arXiv:2604.27542v1 Announce Type: new Abstract: Conventionally, Automatic Speech Recognition (ASR) systems are evaluated on their ability to correctly recognize each word contained in a speech signal.

HAVEN: Hybrid Automated Verification ENgine for UVM Testbench Synthesis with LLMs

ResearchDGX agent

arXiv:2604.27643v1 Announce Type: cross Abstract: Integrated Circuit (IC) verification consumes nearly 70% of the IC development cycle, and recent research leverages Large Language Models (LLMs) to au

HealthBench Professional: Evaluating Large Language Models on Real Clinician Chats

Model ReleasesDGX agent

arXiv:2604.27470v1 Announce Type: new Abstract: Millions of clinicians use ChatGPT to support clinical care, but evaluations of the most common use cases in model-clinician conversations are limited.

HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation

Model ReleasesDGX agent

arXiv:2604.28196v1 Announce Type: new Abstract: Driving world models serve as a pivotal technology for autonomous driving by simulating environmental dynamics. However, existing approaches predominant

Heterogeneous Scientific Foundation Model Collaboration

AgentsDGX agent

arXiv:2604.27351v1 Announce Type: new Abstract: Agentic large language model systems have demonstrated strong capabilities. However, their reliance on language as the universal interface fundamentally

HighFM: Towards a Foundation Model for Learning Representations from High-Frequency Earth Observation Data

Model ReleasesDGX agent

arXiv:2604.04306v2 Announce Type: replace-cross Abstract: The increasing frequency and severity of climate related disasters have intensified the need for real time monitoring, early warning, and info

HiMix: Hierarchical Artifact-aware Mixup for Generalized Synthetic Image Detection

ResearchDGX agent

arXiv:2604.27903v1 Announce Type: new Abstract: The rapid evolution of generative models has enabled the creation of highly realistic and diverse synthetic images, posing significant challenges to rel

Hinge Regression Tree: A Newton Method for Oblique Regression Tree Splitting

Local AiDGX agent

arXiv:2602.05371v3 Announce Type: replace Abstract: Oblique decision trees combine the transparency of trees with the power of multivariate decision boundaries, but learning high-quality oblique split

How Generative AI Disrupts Search: An Empirical Study of Google Search, Gemini, and AI Overviews

Model ReleasesDGX agent

arXiv:2604.27790v1 Announce Type: cross Abstract: Generative AI is being increasingly integrated into web search for the convenience it provides users. In this work, we aim to understand how generativ

How Hard Is Continuous Clustering? Lower Bounds from the Existential Theory of the Reals

SafetyDGX agent

arXiv:2604.26972v1 Announce Type: cross Abstract: This paper studies the computational difficulty of clustering problems that are defined directly on a continuous probability density. Rather than work

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance

SafetyDGX agent

arXiv:2604.27147v1 Announce Type: cross Abstract: In generative modeling, we often wish to produce samples that maximize a user-specified reward such as aesthetic quality or alignment with human prefe

HQ-UNet: A Hybrid Quantum-Classical U-Net with a Quantum Bottleneck for Remote Sensing Image Segmentation

Model ReleasesDGX agent

arXiv:2604.27206v1 Announce Type: new Abstract: Semantic segmentation in remote sensing is commonly addressed using classical deep learning architectures such as U-Net, which require a large number of

Hypencoder Revisited: Reproducibility and Analysis of Non-Linear Scoring for First-Stage Retrieval

ResearchDGX agent

arXiv:2604.27037v1 Announce Type: cross Abstract: The Hypencoder, proposed by Killingback et al., is a retrieval framework that replaces the fixed inner-product scoring function used in standard bi-en

Hyper-Dimensional Fingerprints as Molecular Representations

ResearchDGX agent

arXiv:2604.27810v1 Announce Type: new Abstract: Computational molecular representations underpin virtual screening, property prediction, and materials discovery. Conventional fingerprints are efficien

Hyperspectral Image Classification via Efficient Global Spectral Supertoken Clustering

ResearchDGX agent

arXiv:2604.27364v1 Announce Type: new Abstract: Hyperspectral image classification demands spatially coherent predictions and precise boundary delineation. Yet prevailing superpixel-based methods face

Hypnopaedia-Aware Machine Unlearning via Psychometrics of Artificial Mental Imagery

ResearchDGX agent

arXiv:2410.05284v2 Announce Type: replace-cross Abstract: Neural backdoors represent insidious cybersecurity loopholes that render learning machinery vulnerable to unauthorised manipulations, potentia

IACDM: Interactive Adversarial Convergence Development Methodology -- A Structured Framework for AI-Assisted Software Development

ApplicationsDGX agent

arXiv:2604.16399v2 Announce Type: replace-cross Abstract: The widespread adoption of AI-assisted development tools in 2025 -- and the emergence of vibe coding, a practice of generating complete applic

IKSPARK: Obstacle-Aware Inverse Kinematics via Convex Optimization

ResearchDGX agent

arXiv:2403.12235v2 Announce Type: replace Abstract: Inverse kinematics (IK) is central to robot control and motion planning, yet its nonlinear kinematic mapping makes it inherently nonconvex and parti

ImagineNav++: Prompting Vision-Language Models as Embodied Navigator through Scene Imagination

AgentsDGX agent

arXiv:2512.17435v3 Announce Type: replace Abstract: Visual navigation is a fundamental capability for autonomous home-assistance robots, enabling long-horizon tasks such as object search. While recent

Imitation Game for Adversarial Disillusion with Chain-of-Thought Reasoning in Generative AI

AgentsDGX agent

arXiv:2501.19143v2 Announce Type: replace Abstract: As the cornerstone of artificial intelligence, machine perception confronts a fundamental threat posed by adversarial illusions. These adversarial a

Implicit bias produces neural scaling laws in learning curves, from perceptrons to deep networks

SafetyDGX agent

arXiv:2505.13230v3 Announce Type: replace Abstract: Scaling laws in deep learning -- empirical power-law relationships linking model performance to resource growth -- have emerged as simple yet striki

Improving Calibration in Test-Time Prompt Tuning for Vision-Language Models via Data-Free Flatness-Aware Prompt Pretraining

ApplicationsDGX agent

arXiv:2604.27715v1 Announce Type: new Abstract: Test-time prompt tuning (TPT) has emerged as a promising technique for enhancing the adaptability of vision-language models by optimizing textual prompt

Improving Graph Few-shot Learning with Hyperbolic Space and Denoising Diffusion

Model ReleasesDGX agent

arXiv:2604.27462v1 Announce Type: cross Abstract: Graph few-shot learning, which focuses on effectively learning from only a small number of labeled nodes to quickly adapt to new tasks, has garnered s

In-Context Examples Suppress Scientific Knowledge Recall in LLMs

ResearchDGX agent

arXiv:2604.27540v1 Announce Type: new Abstract: Scientific reasoning rarely stops at what is directly observable; it often requires uncovering hidden structure from data. From estimating reaction cons

In-context Learning vs. Instruction Tuning: The Case of Small and Multilingual Language Models

SafetyDGX agent

arXiv:2503.01611v3 Announce Type: replace Abstract: Instruction following is a critical ability for Large Language Models to perform downstream tasks. The standard approach to instruction tuning has r

In-Context Prompting Obsoletes Agent Orchestration for Procedural Tasks

AgentsDGX agent

arXiv:2604.27891v1 Announce Type: new Abstract: Agent orchestration frameworks -- LangGraph, CrewAI, Google ADK, OpenAI Agents SDK, and others -- place an external orchestrator above the LLM, tracking

In Line with Context: Repository-Level Code Generation via Context Inlining

ResearchDGX agent

arXiv:2601.00376v2 Announce Type: replace-cross Abstract: Repository-level code generation has attracted growing attention in recent years. Unlike function-level code generation, it requires the model

Instruction Complexity Induces Positional Collapse in Adversarial LLM Evaluation

Model ReleasesDGX agent

arXiv:2604.27249v1 Announce Type: cross Abstract: When instructed to underperform on multiple-choice evaluations, do language models engage with question content or fall back on positional shortcuts?

Instruction-Guided Poetry Generation in Arabic and Its Dialects

ResearchDGX agent

arXiv:2604.27766v1 Announce Type: cross Abstract: Poetry has long been a central art form for Arabic speakers, serving as a powerful medium of expression and cultural identity. While modern Arabic spe

Intent2Tx: Benchmarking LLMs for Translating Natural Language Intents into Ethereum Transactions

Model ReleasesDGX agent

arXiv:2604.27763v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) offers a transformative interface for Web3, yet existing benchmarks fail to capture the complexity of tran

Interaction Forces and Internal Loads in Parallel Manipulators with Actuation Redundancy

ApplicationsDGX agent

arXiv:2604.27095v1 Announce Type: new Abstract: This paper discusses null-space wrench components in parallel manipulators. We examine the adaptation of the two most common characterizations of these

InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?

Model ReleasesDGX agent

arXiv:2604.27419v1 Announce Type: new Abstract: With the advancement of multimodal large language models (MLLMs) and coding agents, the website development has shifted from manual programming to agent

Intern-Atlas: A Methodological Evolution Graph as Research Infrastructure for AI Scientists

SafetyDGX agent

arXiv:2604.28158v1 Announce Type: new Abstract: Existing research infrastructure is fundamentally document-centric, providing citation links between papers but lacking explicit representations of meth

InterPartAbility: Text-Guided Part Matching for Interpretable Person Re-Identification

ResearchDGX agent

arXiv:2604.27122v1 Announce Type: new Abstract: Text-to-image person re-identification (TI-ReID) relies on natural-language text description to retrieve top matching individuals from a large gallery o

Interval Orders, Biorders and Credibility-limited Belief Revision

AgentsDGX agent

arXiv:2604.27156v1 Announce Type: new Abstract: Rational belief revision is commonly viewed as being based on a preference order between possible worlds, with the resulting new belief set being those

Investigating More Explainable and Partition-Free Compositionality Estimation for LLMs: A Rule-Generation Perspective

ResearchDGX agent

arXiv:2604.27340v1 Announce Type: new Abstract: Compositional generalization tests are often used to estimate the compositionality of LLMs. However, such tests have the following limitations: (1) they

Iterative Definition Refinement for Zero-Shot Classification via LLM-Based Semantic Prototype Optimization

Model ReleasesDGX agent

arXiv:2604.27335v1 Announce Type: new Abstract: Web filtering systems rely on accurate web content classification to block cyber threats, prevent data exfiltration, and ensure compliance. However, cla

Iterative Multimodal Retrieval-Augmented Generation for Medical Question Answering

Model ReleasesDGX agent

arXiv:2604.27724v1 Announce Type: new Abstract: Medical retrieval-augmented generation (RAG) systems typically operate on text chunks extracted from biomedical literature, discarding the rich visual c

ITS-Mina: A Harris Hawks Optimization-Based All-MLP Framework with Iterative Refinement and External Attention for Multivariate Time Series Forecasting

Model ReleasesDGX agent

arXiv:2604.27981v1 Announce Type: cross Abstract: Multivariate time series forecasting plays a pivotal role in numerous real-world applications, including financial analysis, energy management, and tr

JaiTTS: A Thai Voice Cloning Model

ApplicationsDGX agent

arXiv:2604.27607v1 Announce Type: new Abstract: We present JaiTTS-v1.0, a state-of-the-art Thai voice cloning text-to-speech model built through continual training on a large Thai-centric speech corpu

JI-ADF: Joint-Individual Learning with Adaptive Decision Fusion for Multimodal Skin Lesion Classification

Model ReleasesDGX agent

arXiv:2604.27343v1 Announce Type: new Abstract: Skin lesion classification is essential for early dermatological diagnosis, yet many existing computer-aided systems rely primarily on dermoscopic image

Judge, Then Drive: A Critic-Centric Vision Language Action Framework for Autonomous Driving

Model ReleasesDGX agent

arXiv:2604.27366v1 Announce Type: new Abstract: Recent advances in vision language action (VLA) models have shown remarkable potential for autonomous driving by directly mapping multimodal inputs to c

Junk DNA Hypothesis: Pruning Small Pre-Trained Weights Irreversibly and Monotonically Impairs 'Difficult' Downstream Tasks in LLMs

ResearchDGX agent

arXiv:2310.02277v4 Announce Type: replace-cross Abstract: We present Junk DNA Hypothesis by adopting a novel task-centric angle for the pre-trained weights of large language models (LLMs). It has been

K2MUSE: A human lower-limb multimodal walking dataset spanning task and acquisition variability for rehabilitation robotics

Model ReleasesDGX agent

arXiv:2504.14602v2 Announce Type: replace-cross Abstract: The natural interaction and control performance of lower limb rehabilitation robots are closely linked to biomechanical information from vario

KellyBench: A Benchmark for Long-Horizon Sequential Decision Making

Model ReleasesDGX agent

arXiv:2604.27865v1 Announce Type: new Abstract: Language models are saturating benchmarks for procedural tasks with narrow objectives. But they are increasingly being deployed in long-horizon, non-sta

Kernelized Advantage Estimation: From Nonparametric Statistics to LLM Reasoning

SafetyDGX agent

arXiv:2604.28005v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have increasingly relied on reinforcement learning (RL) to improve their reasoning capabilities. Three a

Knowledge Affordances for Hybrid Human-AI Information Seeking

AgentsDGX agent

arXiv:2604.27539v1 Announce Type: cross Abstract: As information ecosystems grow more heterogeneous, both humans and artificial agents increasingly face a simple yet unresolved question: when seeking

Knowledge Graph Representations for LLM-Based Policy Compliance Reasoning

SafetyDGX agent

arXiv:2604.27713v1 Announce Type: new Abstract: The risks posed by AI features are increasing as they are rapidly integrated into software applications. In response, regulations and standards for safe

LA-Pose: Latent Action Pretraining Meets Pose Estimation

SafetyDGX agent

arXiv:2604.27448v1 Announce Type: new Abstract: This paper revisits camera pose estimation through the lens of self-supervised pretraining, focusing on inverse-dynamics pretraining as a scalable alter

Language Ideologies in a Multilingual Society: An LLM-based Analysis of Luxembourgish News Comments

ResearchDGX agent

arXiv:2604.27661v1 Announce Type: new Abstract: Detecting language ideologies is a valuable yet complex task for understanding how identities are constructed through discourse. In Luxembourg's multicu

Language Models Refine Mechanical Linkage Designs Through Symbolic Reflection and Modular Optimisation

Model ReleasesDGX agent

arXiv:2604.27962v1 Announce Type: new Abstract: Designing mechanical linkages involves combinatorial topology selection and continuous parameter fitting. We show that language models can systematicall

LaST-R1: Reinforcing Action via Adaptive Physical Latent Reasoning for VLA Models

Model ReleasesDGX agent

arXiv:2604.28192v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have increasingly incorporated reasoning mechanisms for complex robotic manipulation. However, existing approaches

Latent Adversarial Detection: Adaptive Probing of LLM Activations for Multi-Turn Attack Detection

ApplicationsDGX agent

arXiv:2604.28129v1 Announce Type: cross Abstract: Multi-turn prompt injection follows a known attack path -- trust-building, pivoting, escalation but text-level defenses miss covert attacks where indi

Latent-GRPO: Group Relative Policy Optimization for Latent Reasoning

SafetyDGX agent

arXiv:2604.27998v1 Announce Type: cross Abstract: Latent reasoning offers a more efficient alternative to explicit reasoning by compressing intermediate reasoning into continuous representations and s

Leading Across the Spectrum of Human-AI Relationships: A Conceptual Framework for Increasingly Heterogeneous Teams

ResearchDGX agent

arXiv:2604.27392v1 Announce Type: new Abstract: What shapes a consequential decision when human and artificial intelligence work on it together? The answer is becoming harder to see. A decision may lo

Learning-Based Hierarchical Scene Graph Matching for Robot Localization Leveraging Prior Maps

AgentsDGX agent

arXiv:2604.27821v1 Announce Type: new Abstract: Accurate localization is a fundamental requirement for autonomous robots operating in indoor environments. Scene graphs encode the spatial structure of

Learning from a single labeled face and a stream of unlabeled data

ResearchDGX agent

arXiv:2604.27564v1 Announce Type: new Abstract: Face recognition from a single image per person is a challenging problem because the training sample is extremely small. We consider a variation of this

← Previous
1…800801802803804…998
Next →