AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
14 Aug 2026

LLMs Know the Constraint But Do Not Use It: Activation Bottlenecks in Pragmatic Constraint Reasoning

SafetyDGX agent

arXiv:2608.12321v1 Announce Type: cross Abstract: When a salient surface cue competes with an implicit feasibility constraint, LLMs often fail -- but aggregate accuracy conflates genuine constraint in

Momentum as Residual-Driven Multiplier Correction for Deep Learning Optimization

ResearchDGX agent

arXiv:2608.12925v1 Announce Type: new Abstract: Momentum-based optimizers are widely used in modern deep learning, yet the relations among momentum recursion, update geometry, and acceleration remain

Moose: Latent concept learning with reasoning-shortcut awareness in EL^{++}

ApplicationsDGX agent

arXiv:2608.12961v1 Announce Type: new Abstract: The OWL 2 EL profile is used in some of the largest production ontologies, including the Gene Ontology and SNOMED CT. Existing neuro-symbolic (NeSy) lea

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

On the global feature importance for interpretable and trustworthy heat demand forecasting

SafetyDGX agent

arXiv:2608.13039v1 Announce Type: new Abstract: The paper introduces the ante-hoc Explainable AI methodology to assess the global feature importance of the Machine Learning models used for heat demand

P2Fusion: Prompt-based Progressive Infrared-Visible Image Fusion via Dual-Prior Distillation

ResearchDGX agent

arXiv:2608.13045v1 Announce Type: new Abstract: Infrared-visible image fusion (IVIF) is pivotal for multimodal perception, yet reconciling the inherent information disparity between thermal and textur

PatientAct: Theory-Grounded Mental Health Client Simulation

ResearchDGX agent

arXiv:2608.12750v1 Announce Type: cross Abstract: LLM-based simulated clients are increasingly used to train novice counselors, evaluate LLM therapists, and generate synthetic data. However, current s

Physics-informed distribution of relaxation times estimation and latent-space condition monitoring of solid oxide fuel and electrolysis cells from electrochemical impedance spectroscopy

ResearchDGX agent

arXiv:2608.13305v1 Announce Type: cross Abstract: Estimating the distribution of relaxation times (DRT) fromelectrochemical impedance spectroscopy (EIS) is an ill-posed inverse problem that is highly

Physics-Informed Laplace Neural Operator for Solving Partial Differential Equations

ResearchDGX agent

arXiv:2602.12706v2 Announce Type: replace Abstract: Neural operators have emerged as fast surrogate solvers for parametric partial differential equations (PDEs). However, purely data-driven models oft

PixSDS: Why Latent SDS Makes Noisy Pixels

ResearchDGX agent

arXiv:2608.12997v1 Announce Type: new Abstract: Score Distillation Sampling (SDS) enables text-to-3D generation by optimizing rendered images with a pretrained diffusion prior, but latent SDS often pr

RoutePack: Expert Placement and Attention-Aware Data Packing for MoE Reinforcement Learning

ResearchDGX agent

arXiv:2608.12146v1 Announce Type: cross Abstract: Training Mixture-of-Experts (MoE) models for reinforcement learning (RL) couples two load-balancing problems: sequence composition determines dense at

RulerNet: Learning Perspective-Invariant Ruler Representations for Robust Image Scale Estimation

Local AiDGX agent

arXiv:2507.07077v2 Announce Type: replace Abstract: Accurately converting pixel measurements into absolute real-world dimensions remains a fundamental challenge in computer vision, limiting progress i

SDS-LoRA: Overcoming Anisotropic Gradient Scaling in Low-Rank Adaptation

SafetyDGX agent

arXiv:2606.16454v2 Announce Type: replace-cross Abstract: Low-Rank Adaptation (LoRA) enables efficient adaptation of large pretrained models to downstream tasks by parameterizing weight updates with l

Spatially-Grounded Text-to-Video Generation via Inference-Time Gradient-Free Optimization

ResearchDGX agent

arXiv:2608.13037v1 Announce Type: new Abstract: Diffusion Transformer Text-to-Video models have achieved remarkable synthesis quality, yet fine-grained spatial controllability remains a significant ch

Taking Qwen3.5-9B quants to SOTA. New lineup incoming :)

Local AiDGX agent

Hey Folks, Some of you saw my Muse-Glimmer-30B SOTA line this week. I recently went all in, got a few more techniques hooked up to my pipeline, and re-quantized the 3.5 9B. And oh boy, did it demolish

Towards Context-Aware Clinical Motion Understanding in Daily Living at Home: Freezing of Gait Detection with Egocentric Vision

ResearchDGX agent

arXiv:2608.13283v1 Announce Type: new Abstract: Understanding motion in daily living requires context beyond kinematics, because similar inertial patterns during activities of daily living (ADLs) can

Tracing Provenance and Detecting Tampering with Complementary LLM Watermarks

ResearchDGX agent

arXiv:2608.12713v1 Announce Type: cross Abstract: Watermarking LLM-generated text is an important task for tracing its provenance. Existing LLM watermarks preserve provenance under editing, but this s

VALG: An Agentic System for ML Theory Research

AgentsDGX agent

arXiv:2608.13060v1 Announce Type: new Abstract: Machine learning theory studies learning procedures through mathematical setups in which the data model, training protocol, oracle access, loss, metric,

Who Speaks Matters: Authority-Aware Multi-View RAG over Italian Parliamentary Proceedings

SafetyDGX agent

arXiv:2608.13410v1 Announce Type: new Abstract: Parliamentary proceedings are a primary record of democratic deliberation, yet their volume and fragmentation make multi-perspective access difficult fo

WinCore

Local AiDGX agent

Looking for testers and feedback for WinCore on Windows I’m working on WinCore, an open-source Python library focused on AI and machine-learning workflows on Windows, particularly around Python/PyTorc

13 Aug 2026

A New First-Order Meta-Learning Algorithm with Convergence Guarantees

SafetyDGX agent

arXiv:2409.03682v2 Announce Type: replace Abstract: Learning new tasks by leveraging prior experience is a fundamental trait of intelligent systems. While Model-Agnostic Meta-Learning (MAML) is a lead

BLADE: Better Language Answers through Dialogue and Explanations

TutorialsDGX agent

arXiv:2604.03236v2 Announce Type: replace-cross Abstract: Large language model (LLM)-based educational assistants often provide direct answers offering little incentive for students to explore or enga

Clustered Randomized Smoothing for Stochastic Prediction Functions

SafetyDGX agent

arXiv:2608.12037v1 Announce Type: new Abstract: Modern stochastic predictors can model rich, multi-modal outcome distributions. However, this expressive power comes with challenges in ensuring robust

Curvature-Aware Zeroth-Order Optimization for Memory-Efficient Test-Time Adaptation

Local AiDGX agent

arXiv:2608.12279v1 Announce Type: new Abstract: Test-time adaptation (TTA) aims to enhance the cross-domain performance of pre-trained models by adapting to unlabeled test data. While most existing TT

DCM Bandits: Multiplayer Information Asymmetric Cascading Bandits for Multiple Clicks

AgentsDGX agent

arXiv:2608.11873v1 Announce Type: new Abstract: In this work, we extend the Dependent Click Model (DCM) Bandits to a multiplayer information-asymmetric setting, where multiple agents interact with a s

Dialogue-Aware Video-to-Music Generation Using Public Domain Film Collections

ResearchDGX agent

arXiv:2608.11576v1 Announce Type: cross Abstract: Video-to-music generation has drawn growing interest for its role in conveying the emotion of visual media, including film. Progress in the field, how

DonorRank: Donor Language Selection for Low-Resource Cross-Lingual Speech Recognition

ResearchDGX agent

arXiv:2608.11441v1 Announce Type: new Abstract: Low-resource automatic speech recognition (ASR) commonly relies on cross-lingual transfer, where models are adapted from higher-resource donor languages

Federated Learning for Distributed CNC Tool Wear Prediction

ApplicationsDGX agent

arXiv:2608.11281v1 Announce Type: cross Abstract: Tool wear prediction is an important task in CNC machining, where accurate monitoring of tool condition supports product quality and process reliabili

Harness-IF: Evaluating Instruction Following Across Instruction Surfaces in Coding Agents

AgentsDGX agent

arXiv:2608.11727v1 Announce Type: new Abstract: When a coding agent obeys a rule, it may simply have been going to do that anyway. Existing instruction-following benchmarks cannot tell the difference:

HyperFix: Combinatorial Nonlinear Correction for Task Vector Merging

ResearchDGX agent

arXiv:2608.11499v1 Announce Type: cross Abstract: Task vectors enable model merging without joint retraining. In practice, the subset of task vectors to be merged may vary, but many existing methods u

If you write rules in an AGENTS.md, this one is worth your time. When a coding agent follows your rule, it may have been going to do that an…

AgentsDGX agent

If you write rules in an AGENTS.md, this one is worth your time. When a coding agent follows your rule, it may have been going to do that anyway. Harness-IF separates the two by scoring 256 rules one

Instruction Alignment for Binary Code Representation Learning

SafetyDGX agent

arXiv:2608.11766v1 Announce Type: cross Abstract: Binary code representation learning is a fundamental problem in software security and reverse engineering. Existing methods mainly learn function-leve

Investigating Learner-Aware Design of LLM-Generated Educational Feedback

ResearchDGX agent

arXiv:2602.11650v2 Announce Type: replace Abstract: Although large language models (LLMs) show promise for generating educational feedback, it remains unclear how feedback should be designed (e.g., to

KANResDiff: Learning Local Residual Diffusion via Kolmogorov-Arnold Network for Ambiguous Medical Image Segmentation

Local AiDGX agent

arXiv:2608.11617v1 Announce Type: new Abstract: Ambiguous medical image segmentation aims to provide a series of diverse but plausible segmentation hypotheses. However, existing methods introduce stoc

Learning from Online User Feedback for Shopping Agents

SafetyDGX agent

arXiv:2608.11604v1 Announce Type: new Abstract: Large language model-based shopping agents are increasingly deployed in real-world e-commerce platforms, generating massive amounts of user interaction

LoongReflect: Boosting Long-Horizon Reflection in Search Agents via Global Perspective Distillation

SafetyDGX agent

arXiv:2608.11967v1 Announce Type: cross Abstract: Large language model agents increasingly rely on long-horizon reasoning to solve complex tasks involving planning, tool use, and memory. A critical ca

Low-Interaction-Rank Learning: Unifying Multiplicative Dual-Encoder Heads

TutorialsDGX agent

arXiv:2608.11661v1 Announce Type: cross Abstract: A multiplicative dual-encoder network computes a real-valued output for a pair of inputs as the inner product of their separate encodings. This archit

Map-Det3D: Metric Feed-Forward 3D Reconstruction Prior for Multi-view 3D Object Detection from Streaming Inputs

Local AiDGX agent

arXiv:2608.12179v1 Announce Type: new Abstract: Metric 3D object detection is a core capability for embodied agents, yet most reliable systems lean on depth sensors, trading away cost, power, and inte

One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL

SafetyDGX agent

arXiv:2608.12253v1 Announce Type: cross Abstract: Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that

Poly-Dialectal Neural Machine Translation System for Bangla Regional Dialects

ApplicationsDGX agent

arXiv:2608.12018v1 Announce Type: new Abstract: Regional dialectal variation poses a fundamental challenge to natural language processing (NLP) in Bangla, where over 240 million speakers communicate a

Principal Trait Analysis: Towards Deriving 'Skills' in Human-AI Collaboration

AgentsDGX agent

arXiv:2608.11460v1 Announce Type: new Abstract: Large Language Model-powered agents are increasingly used in the workplace via human-artificial intelligence (AI) collaboration. In this new era of work

Prompt-Driven Exploration

SafetyDGX agent

arXiv:2607.08837v2 Announce Type: replace-cross Abstract: Exploration is essential to RL since a policy cannot improve by repeatedly sampling the behaviors it already prefers. Standard methods inject

Ready Cohorts: Bounding GPU Opportunity and Avoiding Host Round Trips in LLM-Agent Control

HardwareDGX agent

arXiv:2608.12123v1 Announce Type: cross Abstract: LLM-agent services repeatedly execute small deterministic transitions between model and tool calls: route an outcome, update state, and emit the next

Robust Multi-Tier Infant-Centered Audio Understanding with Whisper via Structured Speaker Conditioning

SafetyDGX agent

arXiv:2608.11587v1 Announce Type: cross Abstract: Recent advances in model design and self-supervised audio representations have improved speech and audio understanding, yet infant-centered naturalist

Row-Bot v4.7.0 is live

Local AiDGX agent

This release brings smaller prompts, longer-running conversations, safer remote access, and more reliable model streaming. External tools and skills are now discovered only when relevant. Long chats g

Small Data Explainer -- The impact of small data methods in everyday life

SafetyDGX agent

arXiv:2507.11773v2 Announce Type: replace-cross Abstract: The emergence of breakthrough artificial intelligence (AI) techniques has led to a renewed focus on how small data settings, i.e., settings wi

STAR: A Spatial-Topology Aware Routing Framework for Generalizable 3D Scene Understanding

ResearchDGX agent

arXiv:2608.11699v1 Announce Type: new Abstract: Constructing a unified 3D scene understanding model has long been hindered by the topological discrepancies across sensor modalities. While applying the

TD-VAD: Breaking Visual Dependence in Video Anomaly Detection with Text-Driven Learning

ResearchDGX agent

arXiv:2608.11820v1 Announce Type: new Abstract: Visual data is typically a prerequisite for training existing video anomaly detection (VAD) methods. However, obtaining sufficient annotated anomaly dat

ToolHazard: Scaling Adversarial Environments for Security Evaluation and Alignment of LLM-based Agents

SafetyDGX agent

arXiv:2608.11878v1 Announce Type: cross Abstract: Large language model (LLM) agents integrated with external tools are vulnerable to indirect prompt injections embedded in environmental states. Howeve

Top-down Traffic Scenario Generation via Joint Initial-Goal Diffusion and Trajectory Infilling

AgentsDGX agent

arXiv:2608.11407v1 Announce Type: new Abstract: Robust traffic simulators are crucial for developing and testing autonomous vehicles to reduce the costly, labor-intensive real-world data collection pr

Towards a Formal Definition of Agent Memory: Basis, Span, Optimality, and the Sequential Memory Problem

SafetyDGX agent

arXiv:2608.11654v1 Announce Type: new Abstract: Despite the wide deployment of memory in large-model agents, there is no unified formal account of what a memory is or when it is optimal. This paper ta

Towards the Harness of Embodied Agents

AgentsDGX agent

arXiv:2608.11246v1 Announce Type: new Abstract: The success of coding agents has established the harness as a paradigm: what an agent achieves depends not on the model alone, but on the infrastructure

TradingMoE: Routing the Right Experts in Evolving Markets

ResearchDGX agent

arXiv:2608.11785v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong potential for financial analysis and trading, but direct trading remains challenging because the predicti

Trust Region Constrained Bayesian Optimization with Penalized Constraint Handling

ApplicationsDGX agent

arXiv:2603.24567v2 Announce Type: replace-cross Abstract: Constrained optimization in high-dimensional black-box settings is difficult due to expensive evaluations, the lack of gradient information, a

While waiting for the release of Qwen3.8-27B, let's try to guess what will happen

AgentsDGX agent

They highlighted 3 things on countdown page: VLM, Agentic Improvements, and Think mode. What improvements do you expect? Reply here! Personally, I want to meet a sage who has attained enlightenment. T

12 Aug 2026

A Graph Neural Network--Guided Genetic Algorithm for Physical Internet Supply Chain Optimization under Cost Uncertainty

ResearchDGX agent

arXiv:2608.10245v1 Announce Type: cross Abstract: Inventory and distribution planning in Physical Internet networks requires coordinating factory-hub assignments, factory supply, lateral transshipment

ASR-Roundtrip Evaluation Can Mask Context- and Convention-Dependent Reading Errors in Chinese News TTS

ResearchDGX agent

arXiv:2608.10606v1 Announce Type: new Abstract: ASR-roundtrip evaluation is widely used as a scalable proxy for text-to-speech (TTS) intelligibility, but it can produce false negatives for reading err

BreastMammo and DenseMammo: Benchmarks for Mammography Domain Generalization

ResearchDGX agent

arXiv:2608.10271v1 Announce Type: cross Abstract: Breast density classification is a critical component of breast cancer risk assessment, yet AI models often struggle to generalize across clinical sit

Capturing Uncertainty in Human Motion for Representation Learning in Soccer

ResearchDGX agent

arXiv:2608.11203v1 Announce Type: new Abstract: This paper presents a self-supervised representation learning framework for understanding 3D skeleton-based human motion in soccer, using future motion

Compact Feed-Forward 3D Gaussians via Saliency-Guided Primitive Merging

ResearchDGX agent

arXiv:2608.10712v1 Announce Type: new Abstract: 3D scene reconstruction, modeling, and rendering are highly relevant for numerous tasks, and 3D Gaussian splatting has become a standard choice in this

ConVAWG: A Retrieval-Grounded Framework for Controlled Synthetic Dialogue Generation in Violence Against Women and Girls

ApplicationsDGX agent

arXiv:2608.11200v1 Announce Type: cross Abstract: Synthetic dialogue generation offers a way to study conversational dynamics in sensitive domains where real data are difficult to access, release, or

← Previous
1…730731732733734…1034
Next →