AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
Human
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
6 Aug 2026

MatrAIx: Simulating the World with 8.3 Billion Persona Agents

Model ReleasesDGX agent

arXiv:2608.04205v1 Announce Type: new Abstract: Human evaluation of AI systems and digital products is costly, slow, and difficult to scale. Offline evaluations are more scalable but often abstract aw

MediRec: Enhancing Chinese Medication Recommendation with Explainable Clinical Reasoning

Model ReleasesDGX agent

arXiv:2510.21084v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong potential for clinical decision support through their advanced language understanding and reaso

MemFly: On-the-Fly Memory Optimization via Information Bottleneck

ResearchDGX agent

arXiv:2602.07885v2 Announce Type: replace Abstract: Long-term memory enables large language model agents to tackle complex tasks through historical interactions. However, existing frameworks encounter

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Memorization in Large Language Models in Medicine: Prevalence, Characteristics, and Implications

SafetyDGX agent

arXiv:2509.08604v5 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated significant potential in medicine, with many studies adapting them through continued pre-traini

MERaLiON-GR: Speech Gender Recognition Model for English and SEA Languages

Model ReleasesDGX agent

arXiv:2608.04433v1 Announce Type: cross Abstract: We present MERaLiON-GR, a speech gender recognition system that performs binary classification (female / male) on English and Southeast Asian (SEA) la

MESH: Memory-Efficient Sinkhorn Optimization for Mixture-of-Experts Training

Model ReleasesDGX agent

arXiv:2608.04407v1 Announce Type: cross Abstract: Memory-efficient matrix optimizers such as Sinkhorn gradient descent remove most AdamW optimizer state for dense Transformer matrices, but direct appl

MetaVideoAgent: Automated Video-Agent Evolution for Long-Form Video Understanding

AgentsDGX agent

arXiv:2608.04587v1 Announce Type: new Abstract: Long-form video understanding requires locating sparse, question-relevant evidence in long, multimodal videos. Real-world video distributions differ in

MGSB: Manifold Gated Signature Branch Pressure-Domain Baseline Architecture for Two-Phase Pipeline Flows Under Distributional Shift

SafetyDGX agent

arXiv:2608.04805v1 Announce Type: new Abstract: Leak detection models for multiphase pipelines often degrade when deployed under flow regimes that differ from training. Existing evaluations typically

MIDAS: Multi-LLM Iterative Data-Adaptive Summarization

ApplicationsDGX agent

arXiv:2608.04307v1 Announce Type: cross Abstract: Text summarization is deceptively difficult. While condensing information seems straightforward, real-world enterprise summarization of support ticket

Mimir: A Neuro-Symbolic Memory System with Dynamic Grounding for Embodied Agents in Interactive Environments

Model ReleasesDGX agent

arXiv:2608.04933v1 Announce Type: new Abstract: Long-horizon embodied task requires agents to act under partial observability while preserving both scene belief and execution progress. Flat histories

Mind the Cap: Output-Budget Regimes Change the Measured Multilingual Reasoning Gap

Model ReleasesDGX agent

arXiv:2608.04160v1 Announce Type: new Abstract: Multilingual evaluations report accuracy at a single output-token cap, but languages need different numbers of tokens to express the same content, so th

Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2608.04633v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) methods improve generalization by aligning their representations with 3D scene geometry. However, these methods are

MINT: Tensor Decomposition on Stacked Recurrence Matrices for Time Series Data Mining

ResearchDGX agent

arXiv:2608.04157v1 Announce Type: new Abstract: Recurrence plots are a time series data mining primitive applied to a variety of domains (e.g. star light curves, sound waveforms, CCT telemetry). This

Mixing-Free and Signal-Optimal Learning of Gaussian Graphical Models from Glauber Dynamics

TutorialsDGX agent

arXiv:2607.18559v2 Announce Type: replace-cross Abstract: Gaussian graphical model selection is usually studied under independent sampling, but in many applications the data arise as a single trajecto

MME: Mixture of Mesh Experts with Random Walk Transformer Gating

ResearchDGX agent

arXiv:2603.00828v2 Announce Type: replace Abstract: In recent years, various methods have been proposed for mesh analysis, each offering distinct advantages and often excelling on different object cla

MOAT: Model-Agnostic Randomized Transformations for preventing Efficiency Degradation Attacks on ViTs

ResearchDGX agent

arXiv:2608.04680v1 Announce Type: cross Abstract: To adopt the Vision Transformers (ViTs) in resource-constrained environment, token pruning is widely used to reduce computational cost without impacti

MobileWAM: Bridging World Action Models to Mobile Manipulation with Chain-of-Foresight

Model ReleasesDGX agent

arXiv:2608.04657v1 Announce Type: new Abstract: World action models (WAMs) built on video generation backbones are a rising recipe for robot learning, yet remain confined to tabletop manipulation. Mob

MoCA: Multi-modal Cross-masked Autoencoder for Digital Health Measurements

Model ReleasesDGX agent

arXiv:2506.02260v4 Announce Type: replace-cross Abstract: Wearable devices enable continuous multi-modal physiological and behavioral monitoring, yet analysis of these data streams faces fundamental c

Modality Agreement- and Conflict-Aware Prototype Hypergraph Learning for Multimodal Intent Understanding

Model ReleasesDGX agent

arXiv:2608.04054v1 Announce Type: cross Abstract: Multimodal intent recognition requires understanding not only what textual, acoustic, and visual signals share, but also how they disagree. Such disag

Monsoon Mayhem to Market Waves: Forecasting Fisheries Resilience in Sri Lanka

ApplicationsDGX agent

arXiv:2608.04023v1 Announce Type: cross Abstract: Sri Lanka's fisheries sector is important for jobs and food supply. Between 2019 and 2025, it faced several major problems at the same time, and how t

Monte Carlo Tree Search for Table-to-Multimodal Report Generation

Model ReleasesDGX agent

arXiv:2608.04071v1 Announce Type: new Abstract: Automatically generating professional multimodal reports comprising both textual analysis and visual charts from structured tabular data is a critical c

MOON3.0: Reasoning-aware Multimodal Representation Learning for E-commerce Product Understanding

Model ReleasesDGX agent

arXiv:2604.00513v3 Announce Type: replace-cross Abstract: With the rapid growth of e-commerce, exploring general representations rather than task-specific ones has attracted increasing attention. Alth

Multi-Objective Ranking for Live-Streaming: Balancing Fresh and Delayed Signals with Segment-Aware Targeting

SafetyDGX agent

arXiv:2608.04455v1 Announce Type: cross Abstract: One of the most challenging problems entertainment live-streaming services face in recommendation systems is that user behaviors are sparse and delaye

Multi-View Face and Gesture Animation with Dynamic Gaussians

ResearchDGX agent

arXiv:2608.04722v1 Announce Type: new Abstract: Creating photorealistic 3D human avatars with realistic upper-body motion remains challenging. Existing approaches either focus on the head and overlook

Multicalibration Yields Better Matchings

SafetyDGX agent

arXiv:2511.11413v2 Announce Type: replace Abstract: Consider the problem of finding the best matching in a weighted graph where we only have access to predictions of the actual stochastic weights, bas

Multimodal Alignment Through Joint Kernel Entropic Gromov--Wasserstein Optimal Transport

SafetyDGX agent

arXiv:2608.04234v1 Announce Type: cross Abstract: We study the problem of aligning data from multiple modalities into a shared representation space, focusing on settings where strong pretrained unimod

Multimodal Spatiotemporal Atmospheric Data Assimilation with Latent Flow-matching

ResearchDGX agent

arXiv:2608.05103v1 Announce Type: new Abstract: Data assimilation (DA) uses Bayesian inference to update the state of a numerical forecast model with observed data. In this study, we propose a fundame

MultiPathFormer: Towards a Foundation Model for Multipath Wireless Propagation

Local AiDGX agent

arXiv:2608.05076v1 Announce Type: cross Abstract: Recent advances in machine learning have enabled training of wireless foundation models, which aim to support tasks such as channel estimation, beam p

muSync-GS: Physics-Synchronized Driving Video Synthesis for Weather and Geometric Road Hazards

SafetyDGX agent

arXiv:2608.04412v1 Announce Type: new Abstract: High-quality driving data are essential for autonomous-driving systems and generative world models. However, rare and safety-critical scenarios involvin

MVTOP: Multi-View Transformer-based Object Pose-Estimation

ResearchDGX agent

arXiv:2508.03243v2 Announce Type: replace Abstract: We present MVTOP, a novel transformer-based method for multi-view rigid object pose estimation. Through an early fusion of the view-specific feature

Neighborhood-Aware Dual Biomedical Entity Linking

ResearchDGX agent

arXiv:2608.04144v1 Announce Type: cross Abstract: Biomedical entity linking grounds mentions in clinical and scientific text to entities in a curated knowledge base (KB) with ontological structure, wh

NeuMoSync: End-to-End Neuromodulatory Control for Plasticity and Adaptability in Continual Learning

TutorialsDGX agent

arXiv:2608.04358v1 Announce Type: new Abstract: Continual learning (CL) requires models to learn tasks sequentially, yet deep neural networks often suffer from plasticity loss and poor knowledge trans

Neural Diversity Regularizes Hallucinations in Language Models

Model ReleasesDGX agent

arXiv:2510.20690v3 Announce Type: replace-cross Abstract: Language models continue to hallucinate despite increases in parameters, compute, and data. We propose neural diversity -- decorrelated parall

Neurocomputational Mechanisms of Syntactic Transfer in Bilingual Sentence Production

ApplicationsDGX agent

arXiv:2601.18056v2 Announce Type: replace Abstract: We discuss the benefits of incorporating oscillatory neural mechanisms into the study of bilingual production errors and their traditionally documen

NeuroPB: Scaling Neural Decoding with Pretrained Behavioral Representations

ResearchDGX agent

arXiv:2608.04389v1 Announce Type: new Abstract: Decoding continuous motor trajectories from neural activity is essential for developing practical brain-computer interfaces (BCIs). However, current neu

NodeJEPA: Structure-Conditioned Latent Prediction for Node-Level Graph Self-Supervised Learning

TutorialsDGX agent

arXiv:2608.04381v1 Announce Type: cross Abstract: Self-supervised learning on graphs is largely shaped by contrastive methods that depend on carefully designed augmentations, and by generative methods

NOLLI: A Difficulty-Calibrated Puzzle Benchmark for Diagnosing the English-Korean Performance Gap

Model ReleasesDGX agent

arXiv:2608.04397v1 Announce Type: new Abstract: We introduce NOLLI, a procedurally generated English-Korean puzzle benchmark designed to diagnose where Korean performance gaps arise. It comprises 15 p

Non-asymptotic implicit bias of logistic regression at early-stage gradient descent dynamics

Model ReleasesDGX agent

arXiv:2608.04382v1 Announce Type: new Abstract: Gradient descent has been of particular interest in modern machine learning beyond sole focus on optimization. Implicit bias emerging from optimization,

Non-Stationary Inventory Control with Lead Times

SafetyDGX agent

arXiv:2602.05799v2 Announce Type: replace-cross Abstract: We study non-stationary single-item, periodic-review inventory control problems in which the demand distribution is unknown and may change ove

Nonparametric Goodness-of-fit Testing under Covariate Shift

ResearchDGX agent

arXiv:2608.04860v1 Announce Type: cross Abstract: This paper develops procedures for nonparametric goodness-of-fit testing under covariate shift, where labelled data are drawn from a source population

Not All Redundant Tokens Are Alike: Analyzing Visual Token Pruning through Token Roles

SafetyDGX agent

arXiv:2608.04483v1 Announce Type: new Abstract: Vision-language models (VLMs) process an image as a sequence of visual tokens, which creates a substantial computational bottleneck during inference. Re

Not Every Divergence Should Be Suppressed: Counterfactual Recoverability in On-Policy Distillation

SafetyDGX agent

arXiv:2608.04408v1 Announce Type: cross Abstract: On-policy distillation (OPD) supervises student-visited trajectories, yet divergence-based rules cannot determine whether an erroneous prefix remains

Not Truly Multilingual: Script Consistency as a Missing Dimension in VLM Evaluation

Model ReleasesDGX agent

arXiv:2606.17188v3 Announce Type: replace-cross Abstract: Current multilingual evaluations for Vision-Language Models (VLMs) assume a one-to-one mapping between language and orthography, overlooking b

NSF-HRPT: Neural Semantic Field meets Hierarchical Risk Perception Tree for Safety-Critical Scenario Assessment

SafetyDGX agent

arXiv:2608.04776v1 Announce Type: new Abstract: The ability to accurately assess and anticipate risks in safety-critical scenarios is crucial for autonomous driving systems. While existing research ha

NuclearDiffusion: Text-to-Image Foundation Models for Learning Nuclear Energy Concepts

Model ReleasesDGX agent

arXiv:2608.04030v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) has transformed text-to-image synthesis, yet its ability to represent specialized engineering domains remains

Objects as Audio-Visual Modal Sound Fields

ApplicationsDGX agent

arXiv:2608.05145v1 Announce Type: new Abstract: While modern 3D reconstruction excels at modeling object geometry and appearance, it largely ignores the rich acoustic cues revealed through physical in

OctoLong: Mid-Training On Cross-Repository Code Contexts Enhances Long-Context Modeling

AgentsDGX agent

arXiv:2608.05141v1 Announce Type: new Abstract: Context lengths of language models (LMs) have dramatically increased, driven by the demands for in-context learning, self-improvement, and long-horizon

ODRA: Synthesizing Cognitive Behavioral Therapy Sessions with Structured Chain-Of-Thought and Dynamic Patient Resistance

SafetyDGX agent

arXiv:2608.04524v1 Announce Type: new Abstract: Synthetic generation of Cognitive Behavioral Therapy (CBT) sessions is challenged by two competing demands: adhering to strict therapeutic structure whi

OmniEdit-Bench: A Comprehensive Benchmark for Instruction-based Video Editing

Model ReleasesDGX agent

arXiv:2608.05049v1 Announce Type: new Abstract: Instruction-based video editing (IVE) is an emerging field with broad applications, yet evaluating editing models remains challenging. Existing benchmar

OmniRouting: A Semantic-Coupled Multimodal Benchmark for Constraint-Aware Spatial Reasoning in PCB Routing

Model ReleasesDGX agent

arXiv:2608.04434v1 Announce Type: new Abstract: Recent large language models (LLMs) have demonstrated remarkable progress in constraint-aware navigation, maze reasoning, and graph reasoning. However,

OmniVR: Joint Video-Audio Conditional Generation for Restoring Degraded Historical Films

Model ReleasesDGX agent

arXiv:2608.04224v1 Announce Type: new Abstract: Historical films suffer from co-occurring visual and audio degradations---blur, noise, flicker, hiss, clipping, and dropout---yet existing methods resto

On Hamming-Lipschitz Type Stability of the Subdominant (Minmax) Ultrametric: Theory and Simple Proofs

ResearchDGX agent

arXiv:2608.04014v1 Announce Type: cross Abstract: The subdominant (minmax) ultrametric is a canonical tree-structured summary of a dissimilarity matrix, arising equivalently as the ultrametric induced

On MUON optimization: From non-convergence to an error analysis with Polar Express and the Newton-Schulz polynomial from implementations

ResearchDGX agent

arXiv:2608.04607v1 Announce Type: cross Abstract: Stochastic gradient descent (SGD) optimization methods are the standard instruments for the training of deep neural networks (DNNs). In many relevant

On the Effectiveness of Adaptation Strategies for VLM-Based Federated Learning in Remote Sensing

Model ReleasesDGX agent

arXiv:2608.04791v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative training of deep learning models across decentralized image archives without requiring data centralization

On The Suitability of Differential Dataflow For Datalog Interpretation In Highly Dynamic Settings

ResearchDGX agent

arXiv:2308.04214v2 Announce Type: replace-cross Abstract: In the domain of knowledge representation and reasoning within AI, datalog engines play an ever-increasingly crucial role. The crux of their o

One Surrogate to Fool Them All: Universal, Transferable, and Targeted Adversarial Attacks with CLIP

Model ReleasesDGX agent

arXiv:2505.19840v3 Announce Type: replace-cross Abstract: Deep Neural Networks (DNNs) have achieved widespread success yet remain prone to adversarial attacks. Typically, such attacks either involve f

OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents

AgentsDGX agent

arXiv:2608.05013v1 Announce Type: cross Abstract: LLM agents are increasingly applied to open-ended everyday requests that span work, study, and life. These tasks are long-horizon, cross-environment,

OPD-V: Visual On-Policy Self-Distillation with Modality Balance

SafetyDGX agent

arXiv:2608.05131v1 Announce Type: cross Abstract: On-Policy Self-Distillation (OPSD) has become a standard post-training approach for improving visual reasoning in multimodal large language models (ML

Optimal Constrained sc-LTL Planning in MDPs via Switching Policies

SafetyDGX agent

arXiv:2608.05021v1 Announce Type: new Abstract: We study the synthesis of optimal policies for planning problems on Markov decision processes with both objectives and safety constraints specified in c

Optimal Training-Time Scaling in Gradual Adaptation

ResearchDGX agent

arXiv:2608.04927v1 Announce Type: new Abstract: In gradual adaptation, how should the training time on each task change as the number of intermediate tasks increases? We study this question for overpa

← Previous
1…7172737475…998
Next →