AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
21 Apr 2026

Culture-Aware Humorous Captioning: Multimodal Humor Generation across Cultural Contexts

SafetyDGX agent

arXiv:2604.18091v1 Announce Type: new Abstract: Recent multimodal large language models have shown promising ability in generating humorous captions for images, yet they still lack stable control over

DAG-STL: A Hierarchical Framework for Zero-Shot Trajectory Planning under Signal Temporal Logic Specifications

TutorialsDGX agent

arXiv:2604.18343v1 Announce Type: new Abstract: Signal Temporal Logic (STL) is a powerful language for specifying temporally structured robotic tasks. Planning executable trajectories under STL constr

DreamShot: Personalized Storyboard Synthesis with Video Diffusion Prior

SafetyDGX agent

arXiv:2604.17195v1 Announce Type: new Abstract: Storyboard synthesis plays a crucial role in visual storytelling, aiming to generate coherent shot sequences that visually narrate cinematic events with

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Dual Strategies for Test-Time Adaptation

ResearchDGX agent

arXiv:2604.17542v1 Announce Type: new Abstract: Conventional test-time adaptation (TTA) approaches typically adapt the model using only a small fraction of test samples, often those with low-entropy p

EAST: Early Action Prediction Sampling Strategy with Token Masking

ResearchDGX agent

arXiv:2604.18367v1 Announce Type: new Abstract: Early action prediction seeks to anticipate an action before it fully unfolds, but limited visual evidence makes this task especially challenging. We in

EchoChain: A Full-Duplex Benchmark for State-Update Reasoning Under Interruptions

Model ReleasesDGX agent

arXiv:2604.16456v1 Announce Type: new Abstract: Real-time voice assistants must revise task state when users interrupt mid-response, but existing spoken-dialog benchmarks largely evaluate turn-based i

EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents

Model ReleasesDGX agent

arXiv:2604.18271v1 Announce Type: new Abstract: As the world of agentic artificial intelligence applied to robotics evolves, the need for agents capable of building and retrieving memories and observa

ENTIRE: Learning-based Volume Rendering Time Prediction

Model ReleasesDGX agent

arXiv:2501.12119v3 Announce Type: replace-cross Abstract: We introduce ENTIRE, a novel deep learning-based approach for fast and accurate volume rendering time prediction. Predicting rendering time is

FLARE: A Data-Efficient Surrogate for Predicting Displacement Fields in Directed Energy Deposition

Model ReleasesDGX agent

arXiv:2604.16649v1 Announce Type: new Abstract: Directed energy deposition (DED) produces complex thermo-mechanical responses that can lead to distortion and reduced dimensional accuracy of a manufact

Forest Before Trees: Latent Superposition for Efficient Visual Reasoning

Local AiDGX agent

arXiv:2601.06803v2 Announce Type: replace Abstract: While Chain-of-Thought empowers Large Vision-Language Models with multi-step reasoning, explicit textual rationales suffer from an information bandw

From Fallback to Frontline: When Can LLMs be Superior Annotators of Human Perspectives?

ResearchDGX agent

arXiv:2604.17968v1 Announce Type: cross Abstract: Although large language models (LLMs) are increasingly used as annotators at scale, they are typically treated as a pragmatic fallback rather than a f

From User Recognition to Activity Counting: An Identity-Agnostic Approach to Multi-User WiFi Sensing

ResearchDGX agent

arXiv:2604.16572v1 Announce Type: new Abstract: Wi-Fi Channel State Information (CSI) enables device-free human activity recognition, but existing multi-user approaches assume a fixed set of known use

GenericAgent: A Token-Efficient Self-Evolving LLM Agent via Contextual Information Density Maximization (V1.0)

AgentsDGX agent

arXiv:2604.17091v1 Announce Type: new Abstract: Long-horizon large language model (LLM) agents are fundamentally limited by context. As interactions become longer, tool descriptions, retrieved memorie

Having a 20-agent system is a million times more powerful than 100 agents working in silos. Last night, I stayed up way too late fixing and …

Model ReleasesDGX agent

Having a 20-agent system is a million times more powerful than 100 agents working in silos. Last night, I stayed up way too late fixing and improving my digital workforce for my AI Agent Mastermind pr

HyKey: Hyperspectral Keypoint Detection and Matching in Minimally Invasive Surgery

ResearchDGX agent

arXiv:2604.17446v1 Announce Type: new Abstract: Purpose: 3D reconstruction in minimally invasive surgery (MIS) enables enhanced surgical guidance through improved visualisation, tool tracking, and aug

I agree, and I will add to this. Companies should own their memories, skills and other resources agents need. Use Claude code, codex, copilo…

Model ReleasesDGX agent

I agree, and I will add to this. Companies should own their memories, skills and other resources agents need. Use Claude code, codex, copilot, etc for intelligence but your execution layer should plug

InternScenes: A Large-scale Simulatable Indoor Scene Dataset with Realistic Layouts

Model ReleasesDGX agent

arXiv:2509.10813v3 Announce Type: replace Abstract: The advancement of Embodied AI heavily relies on large-scale, simulatable 3D scene datasets characterized by scene diversity and realistic layouts.

K2.6 + hermes = 4 hr session setting up qwen 3.6 training regime on dgx spark with the current autnomous session currently lasting 70+ min w…

Model ReleasesDGX agent

K2.6 + hermes = 4 hr session setting up qwen 3.6 training regime on dgx spark with the current autnomous session currently lasting 70+ min without any prompting. Kimi with hermes is next level. @NousR

Kimi K2.6 demonstrates strong long-horizon coding in complex engineering tasks: Kimi K2.6 successfully downloaded and deployed the Qwen3.5-0…

Local AiDGX agent

Kimi K2.6 demonstrates strong long-horizon coding in complex engineering tasks: Kimi K2.6 successfully downloaded and deployed the Qwen3.5-0.8B model locally on a Mac. By implementing and optimizing m

LIFT the Veil for the Truth: Principal Weights Emerge after Rank Reduction for Reasoning-Focused Supervised Fine-Tuning

Model ReleasesDGX agent

arXiv:2506.00772v2 Announce Type: replace-cross Abstract: Recent studies have shown that supervised fine-tuning of LLMs on a small number of high-quality datasets can yield strong reasoning capabiliti

LLaMA-XR: A Novel Framework for Radiology Report Generation using LLaMA and QLoRA Fine Tuning

Model ReleasesDGX agent

arXiv:2506.03178v2 Announce Type: replace-cross Abstract: Automated radiology report generation holds significant potential to reduce radiologists' workload and enhance diagnostic accuracy. However, g

LLMs are still not consistent judges of qualitative work, and small changes to how that work is presented affect outcomes. Better harnessing…

Model ReleasesDGX agent

LLMs are still not consistent judges of qualitative work, and small changes to how that work is presented affect outcomes. Better harnessing and methods (multiple judging runs with randomized orders,

LOGICAL-COMMONSENSEQA: A Benchmark for Logical Commonsense Reasoning

Model ReleasesDGX agent

arXiv:2601.16504v3 Announce Type: replace Abstract: Commonsense reasoning often involves evaluating multiple plausible interpretations rather than selecting a single atomic answer, yet most benchmarks

LVLMs and Humans Ground Differently in Referential Communication

ResearchDGX agent

arXiv:2601.19792v3 Announce Type: replace Abstract: For generative AI agents to partner effectively with human users, the ability to accurately predict human intent is critical. But this ability to co

Macaron: Controlled, Human-Written Benchmark for Multilingual and Multicultural Reasoning via Template-Filling

Model ReleasesDGX agent

arXiv:2602.10732v2 Announce Type: replace Abstract: Multilingual benchmarks rarely test reasoning over culturally grounded premises: translated datasets keep English-centric scenarios, while culture-f

MARCO: Navigating the Unseen Space of Semantic Correspondence

Model ReleasesDGX agent

arXiv:2604.18267v1 Announce Type: new Abstract: Recent advances in semantic correspondence rely on dual-encoder architectures, combining DINOv2 with diffusion backbones. While accurate, these billion-

Marrying Text-to-Motion Generation with Skeleton-Based Action Recognition

Model ReleasesDGX agent

arXiv:2604.17090v1 Announce Type: new Abstract: Human action recognition and motion generation are two active research problems in human-centric computer vision, both aiming to align motion with textu

Measuring Distribution Shift in User Prompts and Its Effects on LLM Performance

ApplicationsDGX agent

arXiv:2604.17650v1 Announce Type: new Abstract: LLMs are increasingly deployed in dynamic, real-world settings, where the distribution of user prompts can shift substantially over time as new tasks, p

Medical Image Understanding Improves Survival Prediction via Visual Instruction Tuning

Model ReleasesDGX agent

arXiv:2604.18250v1 Announce Type: new Abstract: Accurate prognostication and risk estimation are essential for guiding clinical decision-making and optimizing patient management. While radiologist-ass

Memorize When Needed: Decoupled Memory Control for Spatially Consistent Long-Horizon Video Generation

TutorialsDGX agent

arXiv:2604.18215v1 Announce Type: new Abstract: Spatially consistent long-horizon video generation aims to maintain temporal and spatial consistency along predefined camera trajectories. Existing meth

MeSH: Memory-as-State-Highways for Recursive Transformers

Model ReleasesDGX agent

arXiv:2510.07739v2 Announce Type: replace Abstract: Recursive transformers reuse parameters and iterate over hidden states multiple times, decoupling compute depth from parameter depth. However, under

MetaMem: Evolving Meta-Memory for Knowledge Utilization through Self-Reflective Symbolic Optimization

TutorialsDGX agent

arXiv:2602.11182v2 Announce Type: replace Abstract: Existing memory systems enable Large Language Models (LLMs) to support long-horizon human-LLM interactions by persisting historical interactions bey

Multilevel neural networks with dual-stage feature fusion for human activity recognition

Model ReleasesDGX agent

arXiv:2604.16577v1 Announce Type: new Abstract: Human activity recognition (HAR) refers to the process of identifying human actions and activities using data collected from sensors. Neural networks, s

Multimodal Fusion of Histopathology Images and Electronic Health Records for Early Breast Cancer Diagnosis

ResearchDGX agent

arXiv:2604.17122v1 Announce Type: new Abstract: Breast cancer is a leading cause of cancer-related mortality worldwide, and timely accurate diagnosis is critical to improving survival outcomes. While

Navigating the Conceptual Multiverse

SafetyDGX agent

arXiv:2604.17815v1 Announce Type: cross Abstract: When language models answer open-ended problems, they implicitly make hidden decisions that shape their outputs, leaving users with uncontextualized a

Noise Injection: Improving Out-of-Distribution Generalization for Limited Size Datasets

TutorialsDGX agent

arXiv:2511.03855v2 Announce Type: replace Abstract: Deep learned (DL) models for image recognition have been shown to fail to generalize to data from different devices, populations, etc. COVID-19 dete

NTIRE 2026 Rip Current Detection and Segmentation (RipDetSeg) Challenge Report

Model ReleasesDGX agent

arXiv:2604.17070v1 Announce Type: new Abstract: This report presents the NTIRE 2026 Rip Current Detection and Segmentation (RipDetSeg) Challenge, which targets automatic rip current understanding in i

On Safety Risks in Experience-Driven Self-Evolving Agents

SafetyDGX agent

arXiv:2604.16968v1 Announce Type: new Abstract: Experience-driven self-evolution has emerged as a promising paradigm for improving the autonomy of large language model agents, yet its reliance on self

Open-TQ-Metal: Fused Compressed-Domain Attention for Long-Context LLM Inference on Apple Silicon

Model ReleasesDGX agent

arXiv:2604.16957v1 Announce Type: new Abstract: We present Open-TQ-Metal, the first implementation of fused compressed-domain attention on Apple Silicon, enabling 128K-context inference for Llama 3.1

Overcoming Selection Bias in Statistical Studies With Amortized Bayesian Inference

Model ReleasesDGX agent

arXiv:2604.18319v1 Announce Type: cross Abstract: Selection bias arises when the probability that an observation enters a dataset depends on variables related to the quantities of interest, leading to

Policy Testing in Markov Decision Processes

SafetyDGX agent

arXiv:2505.15342v2 Announce Type: replace-cross Abstract: We study the policy testing problem in discounted Markov decision processes (MDPs) in the fixed-confidence setting under a generative model wi

Polysemantic Experts, Monosemantic Paths: Routing as Control in MoEs

Model ReleasesDGX agent

arXiv:2604.17837v1 Announce Type: cross Abstract: An LLM's residual stream is both state and instruction: it encodes the current context and determines the next transformation. We introduce a paramete

Probabilistic Programs of Thought

HardwareDGX agent

arXiv:2604.17290v1 Announce Type: new Abstract: LLMs are widely used for code generation and mathematical reasoning tasks where they are required to generate structured output. They either need to rea

Projected Coupled Diffusion for Test-Time Constrained Joint Generation

ResearchDGX agent

arXiv:2508.10531v3 Announce Type: replace Abstract: Modifications to test-time sampling have emerged as an important extension to diffusion algorithms, with the goal of biasing the generative process

Prompt Sensitivity in Vision-Language Grounding: How Small Changes in Wording Affect Object Detection

ResearchDGX agent

arXiv:2604.17126v1 Announce Type: new Abstract: Vision-language models enable open-vocabulary object grounding through natural language queries, under the implicit assumption that semantically equival

RainFusion2.0: Temporal-Spatial Awareness and Hardware-Efficient Block-wise Sparse Attention

HardwareDGX agent

arXiv:2512.24086v2 Announce Type: replace Abstract: In video and image generation tasks, Diffusion Transformer (DiT) models incur extremely high computational costs due to attention mechanisms, which

Random Matrix Theory of Early-Stopped Gradient Flow: A Transient BBP Scenario

ResearchDGX agent

arXiv:2604.18450v1 Announce Type: cross Abstract: Empirical studies of trained models often report a transient regime in which signal is detectable in a finite gradient descent time window before over

Randomized Antipodal Search Done Right for Data Pareto Improvement of LLM Unlearning

ResearchDGX agent

arXiv:2604.16591v1 Announce Type: new Abstract: Large language models (LLMs) sometimes memorize undesirable knowledge, which must be removed after deployment. Prior work on machine unlearning has focu

Reliability-Aware Adaptive Self-Consistency for Efficient Sampling in LLM Reasoning

Model ReleasesDGX agent

arXiv:2601.02970v2 Announce Type: replace Abstract: Self-Consistency improves reasoning reliability through multi-sample aggregation, but incurs substantial inference cost. Adaptive self-consistency m

Rethinking Cross-Dose PET Denoising: Mitigating Averaging Effects via Residual Noise Learning

TutorialsDGX agent

arXiv:2604.16925v1 Announce Type: new Abstract: Cross-dose denoising for low-dose positron emission tomography (LDPET) has been proposed to address the limited generalization of models trained at a si

Reverse Constitutional AI: A Framework for Controllable Toxic Data Generation via Probability-Clamped RLAIF

SafetyDGX agent

arXiv:2604.17769v1 Announce Type: new Abstract: Ensuring the safety of large language models (LLMs) requires robust red teaming, yet the systematic synthesis of high-quality toxic data remains under-e

Revisiting Auxiliary Losses for Conditional Depth Routing: An Empirical Study

Model ReleasesDGX agent

arXiv:2604.17228v1 Announce Type: new Abstract: Conditional depth execution routes a subset of tokens through a lightweight cheap FFN while the remainder execute the standard full FFN at each controll

R&F-Inventory: A Large-Scale Dataset for Monotonic Inventory Estimation in Reach and Frequency Advertising

Model ReleasesDGX agent

arXiv:2604.16821v1 Announce Type: new Abstract: Reach and Frequency (R&F) contract advertising is an important form of widely used brand advertising. Unlike performance advertising, R&F contracts emph

Scalable and Adaptive Parallel Training of Graph Transformer on Large Graphs

HardwareDGX agent

arXiv:2604.16715v1 Announce Type: cross Abstract: Graph foundation models have demonstrated remarkable adaptability across diverse downstream tasks through large-scale pretraining on graphs. However,

Self-Consistency from Only Two Samples: CoT-PoT Ensembling for Efficient LLM Reasoning

ResearchDGX agent

arXiv:2604.17433v1 Announce Type: new Abstract: Self-consistency (SC) is a popular technique for improving the reasoning accuracy of large language models by aggregating multiple sampled outputs, but

Semantically Stable Image Composition Analysisvia Saliency and Gradient Vector Flow Fusion

Model ReleasesDGX agent

arXiv:2604.16500v1 Announce Type: new Abstract: The reliable computational assessment of photographic composition requires features that are discriminative of spatial layout yet robust to semantic con

SIGMA: A Semantic-Grounded Instruction-Driven Generative Multi-Task Recommender at AliExpress

ApplicationsDGX agent

arXiv:2602.22913v2 Announce Type: replace-cross Abstract: With the rapid evolution of Large Language Models (LLMs), generative recommendation is gradually reshaping the paradigm of recommender systems

SpiralFormer: Looped Transformers Can Learn Hierarchical Dependencies via Multi-Resolution Recursion

Model ReleasesDGX agent

arXiv:2602.11698v2 Announce Type: replace Abstract: Recursive (looped) Transformers decouple computational depth from parameter depth by repeatedly applying shared layers, providing an explicit archit

StepPO: Step-Aligned Policy Optimization for Agentic Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.18401v1 Announce Type: new Abstract: General agents have given rise to phenomenal applications such as OpenClaw and Claude Code. As these agent systems (a.k.a. Harnesses) strive for bolder

SynthFix: Adaptive Neuro-Symbolic Code Vulnerability Repair

TutorialsDGX agent

arXiv:2604.17184v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise for automated code repair but often struggle with the complex semantic and structural correctness required.

← Previous
1…492493494495496…1071
Next →