AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
14 Apr 2026

StaMo: Unsupervised Learning of Generalizable Robot Motion from Compact State Representation

SafetyDGX agent

arXiv:2510.05057v2 Announce Type: replace-cross Abstract: A fundamental challenge in embodied intelligence is developing expressive and compact state representations for efficient world modeling and d

Subargument Argumentation Frameworks: Separating Direct Conflict from Structural Dependency

ResearchDGX agent

arXiv:2601.12038v3 Announce Type: replace Abstract: Dung's abstract argumentation frameworks model acceptability solely in terms of an attack relation, thereby conflating two conceptually distinct asp

TAG-Head: Time-Aligned Graph Head for Plug-and-Play Fine-grained Action Recognition

Model ReleasesDGX agent

arXiv:2604.11498v1 Announce Type: new Abstract: Fine-grained human action recognition (FHAR) is challenging because visually similar actions differ by subtle spatio-temporal cues. Many recent systems

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TAPNext++: What's Next for Tracking Any Point (TAP)?

ResearchDGX agent

arXiv:2604.10582v1 Announce Type: new Abstract: Tracking-Any-Point (TAP) models aim to track any point through a video which is a crucial task in AR/XR and robotics applications. The recently introduc

Teaching the Teacher: The Role of Teacher-Student Smoothness Alignment in Genetic Programming-based Symbolic Distillation

SafetyDGX agent

arXiv:2507.22767v3 Announce Type: replace-cross Abstract: Obtaining human-readable symbolic formulas via genetic programming-based symbolic distillation of a deep neural network trained on a target da

Tencent HY-World 2.0 appears to be dropping on April 15 — open-source multimodal 3D world generation from Tencent Hunyuan

Local AiDGX agent

Tencent's HunyuanWorld is an open-source multimodal 3D world generation model from the Tencent Hunyuan team, featuring 360° immersive experiences via panoramic world proxies, mesh export capabilities

TinyGaze: Lightweight Gaze-Gesture Recognition on Commodity Mobile Devices

Model ReleasesDGX agent

arXiv:2604.09658v1 Announce Type: cross Abstract: Gaze gestures can provide hands free input on mobile devices, but practical use requires (i) gestures users can learn and recall and (ii) recognition

Uber CTO Praveen Neppalli Naga says the company's surging use of AI coding tools has maxed out its full-year AI budget just a few months into 2026 (Laura Bratton/The Information)

Model ReleasesDGX agent

Laura Bratton / The Information: Uber CTO Praveen Neppalli Naga says the company's surging use of AI coding tools has maxed out its full-year AI budget just a few months into 2026 — Uber's surging use

UniToolCall: Unifying Tool-Use Representation, Data, and Evaluation for LLM Agents

Model ReleasesDGX agent

arXiv:2604.11557v1 Announce Type: new Abstract: Tool-use capability is a fundamental component of LLM agents, enabling them to interact with external systems through structured function calls. However

VidAudio-Bench: Benchmarking V2A and VT2A Generation across Four Audio Categories

Model ReleasesDGX agent

arXiv:2604.10542v1 Announce Type: cross Abstract: Video-to-Audio (V2A) generation is essential for immersive multimedia experiences, yet its evaluation remains underexplored. Existing benchmarks typic

We benchmarked TranslateGemma against 5 other LLMs on subtitle translation across 6 languages. At first glance the numbers told a clean story, but then human QA added a chapter. [D]

ResearchDGX agent

This r/MachineLearning discussion post details a hands-on benchmark study in which TranslateGemma — Google's open translation model suite built on Gemma 3, available in 4B, 12B, and 27B sizes and cove

WearBCI Dataset: Understanding and Benchmarking Real-World Wearable Brain-Computer Interfaces Signals

Model ReleasesDGX agent

arXiv:2604.09649v1 Announce Type: cross Abstract: Brain-computer interfaces (BCIs) have opened new platforms for human-computer interaction, medical diagnostics, and neurorehabilitation. Wearable BCI

WebForge: Breaking the Realism-Reproducibility-Scalability Trilemma in Browser Agent Benchmark

Model ReleasesDGX agent

arXiv:2604.10988v1 Announce Type: new Abstract: Existing browser agent benchmarks face a fundamental trilemma: real-website benchmarks lack reproducibility due to content drift, controlled environment

When More Thinking Hurts: Overthinking in LLM Test-Time Compute Scaling

ResearchDGX agent

arXiv:2604.10739v1 Announce Type: new Abstract: Scaling test-time compute through extended chains of thought has become a dominant paradigm for improving large language model reasoning. However, exist

13 Apr 2026

A Little Rank Goes a Long Way: Random Scaffolds with LoRA Adapters Are All You Need

Model ReleasesDGX agent

arXiv:2604.08749v1 Announce Type: new Abstract: How many of a neural network's parameters actually encode task-specific information? We investigate this question with LottaLoRA, a training paradigm in

Accelerating Transformer-Based Monocular SLAM via Geometric Utility Scoring

ResearchDGX agent

arXiv:2604.08718v1 Announce Type: cross Abstract: Geometric Foundation Models (GFMs) have recently advanced monocular SLAM by providing robust, calibration-free 3D priors. However, deploying these mod

Adaptive Rigor in AI System Evaluation using Temperature-Controlled Verdict Aggregation via Generalized Power Mean

Model ReleasesDGX agent

arXiv:2604.08595v1 Announce Type: cross Abstract: Existing evaluation methods for LLM-based AI systems, such as LLM-as-a-Judge, verdict systems, and NLI, do not always align well with human assessment

Aligned Agents, Biased Swarm: Measuring Bias Amplification in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2604.08963v1 Announce Type: cross Abstract: While Multi-Agent Systems (MAS) are increasingly deployed for complex workflows, their emergent properties-particularly the accumulation of bias-remai

Artificial intelligence can persuade people to take political actions

ApplicationsDGX agent

arXiv:2604.09200v1 Announce Type: cross Abstract: There is substantial concern about the ability of advanced artificial intelligence to influence people's behaviour. A rapidly growing body of research

Benchmarked @DJLougen ’s Ornstein-27B-v2 Q6_K on my RTX 3090 using hermes-bench, my new open-source benchmarking UI for local LLMs and Herme…

Model ReleasesDGX agent

Benchmarked @DJLougen ’s Ornstein-27B-v2 Q6_K on my RTX 3090 using hermes-bench, my new open-source benchmarking UI for local LLMs and Hermes agents. Ornstein is a Qwen 3.5 27B fine-tune trained on re

Bridging SFT and RL: Dynamic Policy Optimization for Robust Reasoning

SafetyDGX agent

arXiv:2604.08926v1 Announce Type: new Abstract: Post-training paradigms for Large Language Models (LLMs), primarily Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL), face a fundamental dil

Can We Still Hear the Accent? Investigating the Resilience of Native Language Signals in the LLM Era

ResearchDGX agent

arXiv:2604.08568v1 Announce Type: cross Abstract: The evolution of writing assistance tools from machine translation to large language models (LLMs) has changed how researchers write. This study inves

Case-Grounded Evidence Verification: A Framework for Constructing Evidence-Sensitive Supervision

Local AiDGX agent

arXiv:2604.09537v1 Announce Type: cross Abstract: Evidence-grounded reasoning requires more than attaching retrieved text to a prediction: a model should make decisions that depend on whether the prov

CatalogStitch: Dimension-Aware and Occlusion-Preserving Object Compositing for Catalog Image Generation

Model ReleasesDGX agent

arXiv:2604.08836v1 Announce Type: new Abstract: Generative object compositing methods have shown remarkable ability to seamlessly insert objects into scenes. However, when applied to real-world catalo

China has erased the US lead in AI, Stanford HAI’s 2026 AI index reveals

Model ReleasesDGX agent

Stanford University researchers today released their highly anticipated 2026 AI Index Report, revealing a global landscape where artificial intelligence technology is being adopted at record-breaking

ClusterMark: Towards Robust Watermarking for Autoregressive Image Generators with Visual Token Clustering

SafetyDGX agent

arXiv:2508.06656v2 Announce Type: replace Abstract: In-generation watermarking for latent diffusion models has recently shown high robustness in marking generated images for easier detection and attri

Conformal Prediction in Hierarchical Classification with Constrained Representation Complexity

Model ReleasesDGX agent

arXiv:2501.19038v3 Announce Type: replace-cross Abstract: Conformal prediction has emerged as a widely used framework for constructing valid prediction sets in classification and regression tasks. In

Dataset source for AceStep team! XD

Local AiDGX agent

A Reddit post on r/StableDiffusion pointing the AceStep team toward a potential dataset source for training their open-source AI music generation model. AceStep is a music foundation model whose train

Degradation-Robust Fusion: An Efficient Degradation-Aware Diffusion Framework for Multimodal Image Fusion in Arbitrary Degradation Scenarios

TutorialsDGX agent

arXiv:2604.08922v1 Announce Type: new Abstract: Complex degradations like noise, blur, and low resolution are typical challenges in real world image fusion tasks, limiting the performance and practica

Demystifying the Silence of Correctness Bugs in PyTorch Compiler

ResearchDGX agent

arXiv:2604.08720v1 Announce Type: cross Abstract: Performance optimization of AI infrastructure is key to the fast adoption of large language models (LLMs). The PyTorch compiler (torch.compile), a cor

DRBENCHER: Can Your Agent Identify the Entity, Retrieve Its Properties and Do the Math?

Model ReleasesDGX agent

arXiv:2604.09251v1 Announce Type: new Abstract: Deep research agents increasingly interleave web browsing with multi-step computation, yet existing benchmarks evaluate these capabilities in isolation,

Drift-Aware Online Dynamic Learning for Nonstationary Multivariate Time Series: Application to Sintering Quality Prediction

Model ReleasesDGX agent

arXiv:2604.09358v1 Announce Type: new Abstract: Accurate prediction of nonstationary multivariate time series remains a critical challenge in complex industrial systems such as iron ore sintering. In

Dynamic sparsity in tree-structured feed-forward layers at scale

ResearchDGX agent

arXiv:2604.08565v1 Announce Type: cross Abstract: At typical context lengths, the feed-forward MLP block accounts for a large share of a transformer's compute budget, motivating sparse alternatives to

Efficient Spatial-Temporal Focal Adapter with SSM for Temporal Action Detection

ApplicationsDGX agent

arXiv:2604.09164v1 Announce Type: new Abstract: Temporal human action detection aims to identify and localize action segments within untrimmed videos, serving as a pivotal task in video understanding.

EquiformerV3: Scaling Efficient, Expressive, and General SE(3)-Equivariant Graph Attention Transformers

ResearchDGX agent

arXiv:2604.09130v1 Announce Type: cross Abstract: As SE(3)-equivariant graph neural networks mature as a core tool for 3D atomistic modeling, improving their efficiency, expressivity, and physical con

Exploring Cross-lingual Latent Transplantation: Mutual Opportunities and Open Challenges

ResearchDGX agent

arXiv:2412.12686v3 Announce Type: replace Abstract: Current large language models (LLMs) often exhibit imbalances in multilingual capabilities and cultural adaptability, largely attributed to their En

Fast-dVLM: Efficient Block-Diffusion VLM via Direct Conversion from Autoregressive VLM

AgentsDGX agent

arXiv:2604.06832v2 Announce Type: replace Abstract: Vision-language models (VLMs) predominantly rely on autoregressive decoding, which generates tokens one at a time and fundamentally limits inference

FIT-GNN: Faster Inference Time for GNNs that 'FIT' in Memory Using Coarsening

Model ReleasesDGX agent

arXiv:2410.15001v5 Announce Type: replace Abstract: Scalability of Graph Neural Networks (GNNs) remains a significant challenge. To tackle this, methods like coarsening, condensation, and computation

FP8-RL: A Practical and Stable Low-Precision Stack for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2601.18150v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) for large language models (LLMs) is increasingly bottlenecked by rollout (generation), where long output sequence

Generalization and Scaling Laws for Mixture-of-Experts Transformers

Model ReleasesDGX agent

arXiv:2604.09175v1 Announce Type: cross Abstract: We develop a theory of generalization and scaling for Mixture-of-Experts (MoE) Transformers that cleanly separates active per-input capacity from rout

I can't run Ace-Step 1.5 XL on Comfy!?

Local AiDGX agent

This r/StableDiffusion thread addresses user difficulties running ACE-Step 1.5 XL in ComfyUI — an AI music generation model that was released on April 2, 2026, featuring a 4B-parameter DiT decoder for

Identification and Anonymization of Named Entities in Unstructured Information Sources for Use in Social Engineering Detection

ApplicationsDGX agent

arXiv:2604.09016v1 Announce Type: cross Abstract: This study addresses the challenge of creating datasets for cybercrime analysis while complying with the requirements of regulations such as the Gener

In where to use Ollama cloud?

Local AiDGX agent

This Reddit thread on r/ollama likely discusses where and how to use Ollama's cloud offering, which allows models to run without a powerful local GPU by automatically offloading computation to Ollama'

Integrated electro-optic attention nonlinearities for transformers

ResearchDGX agent

arXiv:2604.09512v1 Announce Type: new Abstract: Transformers have emerged as the dominant neural-network architecture, achieving state-of-the-art performance in language processing and computer vision

Learning General Representation of 12-Lead Electrocardiogram with a Joint-Embedding Predictive Architecture

TutorialsDGX agent

arXiv:2410.08559v5 Announce Type: replace-cross Abstract: Electrocardiogram (ECG) captures the heart's electrical signals, offering valuable information for diagnosing cardiac conditions. However, the

Localizing Task Recognition and Task Learning in In-Context Learning via Attention Head Analysis

ResearchDGX agent

arXiv:2509.24164v2 Announce Type: replace Abstract: We investigate the mechanistic underpinnings of in-context learning (ICL) in large language models by reconciling two dominant perspectives: the com

LuMon: A Comprehensive Benchmark and Development Suite with Novel Datasets for Lunar Monocular Depth Estimation

Model ReleasesDGX agent

arXiv:2604.09352v1 Announce Type: new Abstract: Monocular Depth Estimation (MDE) is crucial for autonomous lunar rover navigation using electro-optical cameras. However, deploying terrestrial MDE netw

MATCHA: Efficient Deployment of Deep Neural Networks on Multi-Accelerator Heterogeneous Edge SoCs

Model ReleasesDGX agent

arXiv:2604.09124v1 Announce Type: cross Abstract: Deploying DNNs on System-on-Chips (SoC) with multiple heterogeneous acceleration engines is challenging, and the majority of deployment frameworks can

MedConceal: A Benchmark for Clinical Hidden-Concern Reasoning Under Partial Observability

Model ReleasesDGX agent

arXiv:2604.08788v1 Announce Type: new Abstract: Patient-clinician communication is an asymmetric-information problem: patients often do not disclose fears, misconceptions, or practical barriers unless

MONETA: Multimodal Industry Classification through Geographic Information with Multi Agent Systems

Model ReleasesDGX agent

arXiv:2604.07956v2 Announce Type: replace Abstract: Industry classification schemes are integral parts of public and corporate databases as they classify businesses based on economic activity. Due to

MuTSE: A Human-in-the-Loop Multi-use Text Simplification Evaluator

SafetyDGX agent

arXiv:2604.08947v1 Announce Type: cross Abstract: As Large Language Models (LLMs) become increasingly prevalent in text simplification, systematically evaluating their outputs across diverse prompting

New WAN 2.2 Lightx2v speed lora 260412

Local AiDGX agent

A new speed-focused LoRA for the Wan 2.2 video generation model, released by the LightX2V project on April 26, 2024, shared on the r/StableDiffusion community. The LightX2V distilled LoRA dramatically

PDE-regularized Dynamics-informed Diffusion with Uncertainty-aware Filtering for Long-Horizon Dynamics

ResearchDGX agent

arXiv:2604.09058v1 Announce Type: cross Abstract: Long-horizon spatiotemporal prediction remains a challenging problem due to cumulative errors, noise amplification, and the lack of physical consisten

PerMix-RLVR: Preserving Persona Expressivity under Verifiable-Reward Alignment

SafetyDGX agent

arXiv:2604.08986v1 Announce Type: cross Abstract: Persona prompting has been widely adopted to steer large language models (LLMs) behavior and improve their instruction performance by assigning specif

Physics-Informed Reinforcement Learning of Spatial Density Velocity Potentials for Map-Free Racing

Local AiDGX agent

arXiv:2604.09499v1 Announce Type: new Abstract: Autonomous racing without prebuilt maps is a grand challenge for embedded robotics that requires kinodynamic planning from instantaneous sensor data at

R2G: A Multi-View Circuit Graph Benchmark Suite from RTL to GDSII

Model ReleasesDGX agent

arXiv:2604.08810v1 Announce Type: new Abstract: Graph neural networks (GNNs) are increasingly applied to physical design tasks such as congestion prediction and wirelength estimation, yet progress is

Retrieval Augmented Classification for Confidential Documents

Model ReleasesDGX agent

arXiv:2604.08628v1 Announce Type: cross Abstract: Unauthorized disclosure of confidential documents demands robust, low-leakage classification. In real work environments, there is a lot of inflow and

StaRPO: Stability-Augmented Reinforcement Policy Optimization

SafetyDGX agent

arXiv:2604.08905v1 Announce Type: new Abstract: Reinforcement learning (RL) is effective in enhancing the accuracy of large language models in complex reasoning tasks. Existing RL policy optimization

Task Vectors, Learned Not Extracted: Performance Gains and Mechanistic Insight

ResearchDGX agent

arXiv:2509.24169v2 Announce Type: replace Abstract: Large Language Models (LLMs) can perform new tasks from in-context demonstrations, a phenomenon known as in-context learning (ICL). Recent work sugg

Text-Conditioned Multi-Expert Regression Framework for Fully Automated Multi-Abutment Design

Model ReleasesDGX agent

arXiv:2604.09047v1 Announce Type: new Abstract: Dental implant abutments serve as the geometric and biomechanical interface between the implant fixture and the prosthetic crown, yet their design relie

← Previous
1…498499500501502…1071
Next →