AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
Human
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
58,761 results
12 Aug 2026

LoRCA: LoRA Cycle Adaptation for Histology to HiP-CT Translation with DINOv3

SafetyDGX agent

arXiv:2608.10002v1 Announce Type: cross Abstract: Hierarchical Phase-Contrast Tomography (HiP-CT) is a synchrotron based X-ray imaging technique that enables non-destructive, volumetric imaging of int

Lost in Reconstruction: Aligning Action Representations with Language in Vision-Language-Action Models

ResearchDGX agent

arXiv:2608.10484v1 Announce Type: cross Abstract: Action verbs describe not only the physical outcomes of actions, but also how those actions are performed. Yet action representations in vision-langua

MAD-HOI: Masked Autoregressive Diffusion for Generating Articulated Hand Object Interactions from Text

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.10162v1 Announce Type: new Abstract: Methods for text-based generation of hand-object interaction (HOI) sequences primarily focus on producing smooth, physically plausible trajectories. A t

MammoMix: Leveraging Mixture of Experts for Robust Mammogram Breast Detection

ApplicationsDGX agent

arXiv:2608.10437v1 Announce Type: new Abstract: Breast lesion detection in mammography remains a challenging task due to variations in image quality, lesion appearance, and population demographics acr

MAP-Graph: Provenance-Aware Shared Memory for Multi-Agent Workflows

Model ReleasesDGX agent

arXiv:2608.10509v1 Announce Type: new Abstract: Shared memory helps language-model agents reuse information across long workflows, yet relevant evidence may not be admissible for a particular agent or

Mapping and Measuring the Behavioral Evolution of Large Language Models

Model ReleasesDGX agent

arXiv:2608.11027v1 Announce Type: cross Abstract: Benchmark leaderboards summarize how well a language model performs, but not how its behavior relates to that of other models or changes across genera

MARCO: Click-Intent Decomposition for Calibrated Ads Conversion Prediction

SafetyDGX agent

arXiv:2608.10562v1 Announce Type: new Abstract: Not all clicks are equal. Industrial ads ranking decouples conversion probability into click-through rate (CTR) and post-click conversion rate (CVR), ye

MarkNull: Model-Agnostic Watermark Removal in AI-Generated Images via On-Manifold Latent Manipulation

SafetyDGX agent

arXiv:2608.10166v1 Announce Type: cross Abstract: Digital watermarking has emerged as a critical technique for provenance and copyright attribution in AI-generated imagery, yet its robustness against

Masked Neural Detection for Run-Length-Limited Channel Coding in Molecular Communication

Model ReleasesDGX agent

arXiv:2606.12489v2 Announce Type: replace-cross Abstract: Molecular communication (MC) suffers from severe diffusion memory because molecules released for one symbol may arrive during later symbol int

MazzikaAI: A knowledge-based performance-to-prompt compiler for real-time Arabic maqam accompaniment with a streaming text-to-music model

ApplicationsDGX agent

arXiv:2608.10360v1 Announce Type: cross Abstract: Arabic maqam music microtonal, modal, and built on ornamented call and response is among the traditions most underserved by generative music models, w

MD-ProTector: Positioning Multiple Data-Driven Prototypes for LLM-Generated Text Detection

ResearchDGX agent

arXiv:2608.10459v1 Announce Type: cross Abstract: As LLM-generated content becomes more sophisticated, detection systems for distinguishing those texts from human-written text must operate at scale wh

Measure the Sim-to-Real Gap: Designing an Affordable Real-World Benchmark Platform for Reinforcement Learning in AIoT Systems

Model ReleasesDGX agent

arXiv:2607.10309v2 Announce Type: replace Abstract: Reinforcement learning (RL) is commonly employed to enhance the performance of autonomous systems, including the Autonomous Internet of Things (AIoT

Measuring Semantic Abstractness of SAE Features via Nonlocality

Model ReleasesDGX agent

arXiv:2608.10537v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) have helped uncover mechanistic explanations for LLM behaviours such as reasoning, jailbreaking etc., via understanding the c

MedUP: Awakening Unified Understanding and Perception in Medical Vision-Language Models

SafetyDGX agent

arXiv:2608.10635v1 Announce Type: cross Abstract: Medical Vision-Language Models (Med-VLMs) excel at verbalizing visual content, yet precise visual perception, segmentation, and grounding remain chall

MEGA: Self-Evolving Agent Optimization Infrastructure via Wisdom Graph

AgentsDGX agent

arXiv:2608.10504v1 Announce Type: new Abstract: As coding agents increasingly handle implementation, the central challenge shifts from building individual agents to building an infrastructure that sys

MemSpec: Memory-Aware Runtime for Adaptive Draft Scheduling in Speculative Decoding on Edge Devices

ResearchDGX agent

arXiv:2608.10362v1 Announce Type: cross Abstract: Speculative decoding accelerates autoregressive large language model (LLM) inference by using a lightweight draft model to speculate multiple tokens,

MERA: Model Evolution and Routing with Skill Adaptation for Agentic Systems at Scale

SafetyDGX agent

arXiv:2608.10333v1 Announce Type: new Abstract: LLM agents execute heterogeneous sequences of model calls within a single task: some invocations require careful reasoning, while others are structured

MESA:Task-Adaptive Multi-Structure Evidence Selection for Long-Horizon Agent Memory

AgentsDGX agent

arXiv:2608.10108v1 Announce Type: new Abstract: Long-horizon agents accumulate trajectories spanning hundreds of interleaved reasoning, action, and observation steps, where answering a query may depen

MIDAS: Mutual Information Disentanglement with Uncertainty-Aware Fusion for Incomplete Multimodal Sentiment Analysis

SafetyDGX agent

arXiv:2608.09986v1 Announce Type: new Abstract: Most existing multimodal sentiment analysis approaches assume access to complete multimodal inputs. However, real-world applications frequently encounte

Mind Viruses: Self-Propagating Ideas in Multi-Agent LLM Systems

AgentsDGX agent

arXiv:2608.10218v1 Announce Type: new Abstract: AI agents are becoming more autonomous and increasingly interconnected, exposing them to new emergent risks arising from agent-to-agent interaction. One

MIRA: Medical Image Reflection for Agentic Diagnosis

AgentsDGX agent

arXiv:2608.10827v1 Announce Type: cross Abstract: Medical visual agents can use tools to inspect images and retrieve external knowledge, but indiscriminate tool use may introduce noisy or misleading e

Mitigating Bus Bunching with Reinforcement Learning Enhanced by Semantic Stop Embedding

SafetyDGX agent

arXiv:2608.10207v1 Announce Type: new Abstract: Bus bunching degrades service regularity and increases passenger waiting in high-frequency transit. Existing reinforcement-learning-based holding contro

Mitigating Context Interference for Reliable and Efficient Search Agents

AgentsDGX agent

arXiv:2608.10743v1 Announce Type: new Abstract: Recent research empowers Large Language Models (LLMs) as multi-turn search agents to iteratively retrieve and generate outputs until complex tasks are s

Mixture-of-Experts-based Entropy Model for Learned Image Compression

ResearchDGX agent

arXiv:2608.10947v1 Announce Type: new Abstract: Learned image compression has seen significant progress in recent years with the development of end-to-end learned models that achieve better compressio

MMArt A Multi-Perspective Multimodal Dataset for Visual Art Understanding

ResearchDGX agent

arXiv:2608.10706v1 Announce Type: new Abstract: Recent vision-language models demonstrate impressive general visual understanding, yet their art interpretation remains shallow: they describe surface c

Modelling Geographic Atrophy Progression using Implicit Neural Representations

ResearchDGX agent

arXiv:2608.10807v1 Announce Type: cross Abstract: Age-related Macular Degeneration (AMD) is the major cause of blindness in the Western world. Its late dry phase is characterised by irreversible atrop

MoE Proxy Models for Low-Cost Failure Reproduction and Diagnosis in LLM RL Post-Training

ResearchDGX agent

arXiv:2608.10823v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training of large language models (LLMs) is computationally intensive and involves complex system pipelines with substa

More Accurate, Less Human: Gestalt Grouping in Vision Models

Model ReleasesDGX agent

arXiv:2608.10195v1 Announce Type: new Abstract: Human vision organizes what it sees into wholes: same-colored points group into series, similar marks cohere into categories, and shapes complete into r

Most biomedical publications show signs of LLM-assisted writing

SafetyDGX agent

arXiv:2608.10715v1 Announce Type: cross Abstract: Over the past several years, LLM-powered chatbots and agents have become widely used as a tool for academic writing. LLM-assisted writing can be valua

Motion Artifact-Aware Self-Supervised Representation Learning for 3D Brain MRI Motion Artifact Reduction

ResearchDGX agent

arXiv:2608.10170v1 Announce Type: new Abstract: Patient motion remains a source of image degradation in brain MRI, leading to signal loss, blurring, and geometric distortion that compromise quantitati

MRIComp4Flow: Compression of 3D Brain MRI for Training Multi-Modal Generative Models

TutorialsDGX agent

arXiv:2608.10291v1 Announce Type: cross Abstract: Large-scale multi-modal MRI datasets impose substantial storage and I/O costs, limiting the training of 3D generative models on commodity infrastructu

MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games

AgentsDGX agent

arXiv:2602.24188v2 Announce Type: replace Abstract: We present a scalable and verifiable methodology for evaluating language models in multi-turn interactions, using a suite of collaborative games tha

Multi-Granular Rationale-Guided Molecular LLM for Property Prediction

ResearchDGX agent

arXiv:2608.10480v1 Announce Type: new Abstract: Large language models (LLMs) are widely applied across chemical tasks, such as molecular property prediction, which underpins drug discovery. Molecular

Multi-Level Evidence Aggregation for Robust Facial Phenotype Retrieval in Rare Genetic Disorder Prioritization

ResearchDGX agent

arXiv:2608.11037v1 Announce Type: new Abstract: AI-assisted facial phenotyping supports rare genetic disorder prioritization by retrieving visually similar diagnosed cases from facial image reference

Multi-View Relational Distillation for Spatial Reasoning with Vision-Language Models

SafetyDGX agent

arXiv:2608.10864v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong image and video understanding, yet their visual-spatial representations remain geometrically fragile,

Multiclass Sentiment Analysis for Identifying Political Viewpoints

ResearchDGX agent

arXiv:2608.11049v1 Announce Type: cross Abstract: The rapid growth of social media has created vast amounts of political discourse, which provides valuable opportunities to analyze public opinions and

Multilingual Embedding Probes Fail to Generalize Across Learner Corpora

ResearchDGX agent

arXiv:2604.07095v2 Announce Type: replace Abstract: Do multilingual embedding models encode a language-general representation of proficiency? We investigate this by training linear and non-linear prob

Multimodal Ambivalence and Hesitancy Recognition via Cross-Attention and Gated Fusion

ResearchDGX agent

arXiv:2607.15779v2 Announce Type: replace Abstract: We present a multimodal framework for Ambivalence/Hesitancy (A/H) recognition in video, developed for the ABAW11 challenge at ECCV 2026. The propose

MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment

Local AiDGX agent

arXiv:2608.11167v1 Announce Type: cross Abstract: Existing Multimodal Large Language Models (MLLMs) predominantly rely on image-text pairs for modality alignment pretraining, mapping global image repr

Multimodal Item Parameter Estimation using Simulated Response Probabilitie

Model ReleasesDGX agent

arXiv:2608.10154v1 Announce Type: cross Abstract: We present results from reconstructing multiple-choice model (MCM) and three-parameter logistic (3PL) model curves using a fine-tuned multimodal large

Multiplayer Nash Preference Optimization

SafetyDGX agent

arXiv:2509.23102v4 Announce Type: replace Abstract: Reinforcement learning from human feedback (RLHF) has emerged as the standard paradigm for aligning large language models with human preferences. Ho

Multiple Scale Latents for Learned Image Compression

ResearchDGX agent

arXiv:2608.10952v1 Announce Type: new Abstract: Most learned image compression systems rely on a single latent representation combined with a hyperprior, which limits their ability to efficiently capt

MUSE: A Full-Text Cross-Domain Knowledge Base of Scientific Problems, Solutions, and Rationales

ResearchDGX agent

arXiv:2608.10974v1 Announce Type: new Abstract: Scientific papers contain fine-grained records of problem solving: authors mention technical obstacles and methods that were used to address them, often

MVTrack: Ultrafast Appearance-Free Moving Object Tracking from Compressed Bitstreams

ResearchDGX agent

arXiv:2608.10790v1 Announce Type: cross Abstract: Deploying modern video trackers at scale is bottlenecked by the computational cost of RGB-based object detectors. To this end, we present MVTrack, an

myMediWhisper: Construction of Burmese Medical Speech Corpus and Whisper Fine-Tuning for Clinical Dialogue ASR

Model ReleasesDGX agent

arXiv:2608.11036v1 Announce Type: new Abstract: Although Whisper models benefit from large-scale multilingual pre-training, their performance on Burmese medical speech remains limited. This work prese

Narrative Keyframing for Generative Creative Writing

ResearchDGX agent

arXiv:2608.10337v1 Announce Type: cross Abstract: We introduce narrative keyframing, an interaction technique for AI-assisted creative writing that lets writers specify different types of narrative co

Navigating in Uncertain Environments with Heterogeneous Visibility

ApplicationsDGX agent

arXiv:2603.03495v2 Announce Type: replace Abstract: Navigating an environment with uncertain connectivity requires a strategic balance between minimizing the cost of traversal and seeking information

Navigating the Proximity-Safety Balance: Constraint Decomposition for Human Following in Pedestrian Crowds

SafetyDGX agent

arXiv:2608.10056v1 Announce Type: cross Abstract: Following a target human in crowded environments involves an inherent conflict between staying close to the target and navigating safely among surroun

Navigation Alone Is Not Enough: Evaluating Explanatory Assistive UI Agents

Model ReleasesDGX agent

arXiv:2608.09944v1 Announce Type: cross Abstract: Modern web interfaces are increasingly difficult to use with screen readers, particularly when pages update dynamically or hide important structure be

Neural Introspection Gating for Adaptive KV-Cache Reuse in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2608.10824v1 Announce Type: cross Abstract: Vision-Language-Action(VLA) models map camera images and language instructions directly to motor commands through a single autoregressive transformer.

Neuroevolution Arena: Nested Ecological Evaluation of Update-and-Inheritance Regimes across Neural Architectures

HardwareDGX agent

arXiv:2608.10323v1 Announce Type: new Abstract: Competitive artificial-life systems can rank trained controllers differently under training and ecological evaluation. We present Neuroevolution Arena,

Never Stop Speaking: a Denial-of-Service Attack on End-to-End Speech Language Models

SafetyDGX agent

arXiv:2608.10405v1 Announce Type: cross Abstract: Many studies have shown that specially crafted inputs can induce large language models (LLMs) to generate excessively long outputs, resulting in signi

No Free Labels: Limitations of LLM-as-a-Judge Without Human Grounding

Model ReleasesDGX agent

arXiv:2503.05061v3 Announce Type: replace Abstract: Reliable evaluation of large language models (LLMs) is critical as their deployment rapidly expands, particularly in high-stakes domains such as bus

Nonlinear Model Predictive Control via Sequential Convex Programming for Drone-to-Drone Docking

AgentsDGX agent

arXiv:2608.10542v1 Announce Type: new Abstract: Autonomous mid-air docking of multi-rotor vehicles under disturbance-driven target motion poses a constrained non-linear trajectory optimization challen

Nonlinear multi-study sparse factor analysis

TutorialsDGX agent

arXiv:2601.18128v2 Announce Type: replace-cross Abstract: High-dimensional data often exhibit variation that can be captured by lower-dimensional factors. For high-dimensional data from multiple studi

NullEdit: Stealthy Image Protection via VLM Condition Redirection

Model ReleasesDGX agent

arXiv:2608.10870v1 Announce Type: new Abstract: Modern image editors combine vision-language models (VLMs) with diffusion transformer backbones to modify a single reference image according to instruct

Nutrition Data Infrastructure for the AI Era: Operationalizing FAIR for Agent-Mediated Research

Model ReleasesDGX agent

arXiv:2608.10363v1 Announce Type: new Abstract: AI agents can accelerate nutrition research, but their analyses inherit the identity, semantic, and release ambiguities of the underlying data. We prese

OAA: Three Phases of Vocal Guidance in Human-Drone Teleoperation

TutorialsDGX agent

arXiv:2608.10651v1 Announce Type: new Abstract: Voice-guided teleoperation requires systems that adapt to the evolving dynamics of human guidance. Yet most voice-controlled robot systems treat spoken

Observational Policy Ranking for SMB Financial Guidance from Multi-Action Accounting Logs

SafetyDGX agent

arXiv:2608.10050v1 Announce Type: new Abstract: Small and medium-sized businesses need timely financial guidance, yet historical accounting logs record self-selected and often co-occurring business ch

Off-Axis, On Purpose: Where a Transformer Computes Concepts and Why it Does So

ResearchDGX agent

arXiv:2608.10251v1 Announce Type: new Abstract: A transformer's answer lives on one axis: the direction its unembedding reads. Its intermediate states largely do not, and that off-axis position is usu

← Previous
1…45678…980
Next →