AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,624 results
19 May 2026

Knowledge-to-Verification: Exploring RLVR for LLMs in Knowledge-Intensive Domains

ResearchDGX agent

arXiv:2605.18261v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has demonstrated promising potential to enhance the reasoning capabilities of large language model

Lambda’s NVIDIA HGX 8xB200 on STAC-AI™ LANG6

Model ReleasesDGX agent

What the numbers mean for financial services Executive summary Lambda is the first to publish an audited STAC-AI™ LANG6 result on NVIDIA HGX 8xB200, with independently verified performance data that F

LaPA^2: Length-Aware Prefix and Prompt Attention Augmentation for Long-Form Controllable Text Generation

Model ReleasesDGX agent

arXiv:2508.04047v2 Announce Type: replace Abstract: Prefix-based methods have emerged as a promising paradigm for Controllable Text Generation (CTG) due to their parameter efficiency. However, while e

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Latent Action Reparameterization for Efficient Agent Inference

AgentsDGX agent

arXiv:2605.18597v1 Announce Type: new Abstract: Large language model (LLM) agents often rely on long sequences of low-level textual actions, resulting in large effective decision horizons and high inf

LaunchDarkly launches runtime control layer for the agentic AI era

Model ReleasesDGX agent

LaunchDarkly, a feature control platform that helps developers and software engineers launch and manage products, today announced the launch of AgentControl, a new solution providing real-time managem

Learning in Position-Aware Multinomial Logit Bandits: From Multiplicative to General Position Effects

ResearchDGX agent

arXiv:2605.17238v1 Announce Type: new Abstract: We study the dynamic joint assortment selection and positioning problem, where the attraction of each product depends on both its intrinsic appeal and i

Learning to Reason without External Rewards

SafetyDGX agent

arXiv:2505.19590v5 Announce Type: replace-cross Abstract: Training large language models (LLMs) for complex reasoning via Reinforcement Learning with Verifiable Rewards (RLVR) is effective but limited

LiteFrame: Efficient Vision Encoders Unlock Frame Scaling in Video LLMs

ResearchDGX agent

arXiv:2605.17260v1 Announce Type: new Abstract: The fundamental challenge in scaling Video Large Language Models (Video LLMs) to long-form video lies in managing the explosion of visual-token context

llm-gemini 0.32a0

Model ReleasesDGX agent

I don't have current information about this specific entry, so I'll describe what it likely covers based on the available details. This entry documents the release or update of llm-gemini version 0.32

Lost or Hidden? A Concept-Level Forgetting in Supervised Continual Learning

ResearchDGX agent

arXiv:2605.16374v1 Announce Type: cross Abstract: Continual learning studies how models can adapt to new tasks while retaining previously acquired knowledge. Although a broad spectrum of methods has b

MADP: A Multi-Agent Pipeline for Sustainable Document Processing with Human-in-the-Loop

Model ReleasesDGX agent

arXiv:2605.17159v1 Announce Type: new Abstract: Document processing automation remains a critical challenge in enterprise environments, where traditional manual approaches are labor-intensive and erro

ManiSoft: Towards Vision-Language Manipulation for Soft Continuum Robotics

Model ReleasesDGX agent

arXiv:2605.18617v1 Announce Type: cross Abstract: Most existing vision-language manipulation research targets rigid robotic arms, whose fixed morphology limits adaptability in cluttered or confined sp

MARS: Technical Report for the CASTLE Challenge at EgoVis 2026

Model ReleasesDGX agent

arXiv:2605.18176v1 Announce Type: cross Abstract: This report presents MARS, short for Multimodal Agentic Reasoning with Source selection, our system for the CASTLE Challenge at EgoVis 2026. Participa

MaskAttn-SDXL: Controllable Region-Level Text-To-Image Generation

Local AiDGX agent

arXiv:2509.15357v2 Announce Type: replace Abstract: Diffusion models have achieved strong results in text-to-image generation, but important limitations remain as prompts become more structured and mu

Mitigating Conversational Inertia in Multi-Turn Agents

SafetyDGX agent

arXiv:2602.03664v3 Announce Type: replace Abstract: Large language models excel as few-shot learners when provided with appropriate demonstrations, yet this strength becomes problematic in multiturn a

MorphSeek: Fine-grained Latent Representation-Level Policy Optimization for Deformable Image Registration

Model ReleasesDGX agent

arXiv:2511.17392v3 Announce Type: replace Abstract: Deformable image registration (DIR) remains a fundamental yet challenging problem in medical image analysis, largely due to the prohibitively high-d

Multi-Mode Quantum Annealing for Generative Representation Learning with Boltzmann Priors

ResearchDGX agent

arXiv:2604.00919v2 Announce Type: replace-cross Abstract: Energy-based models provide a natural bridge between statistical physics and machine learning by representing data through structured energy l

Multi-Object Tracking Consistently Improves Wildlife Inference

ApplicationsDGX agent

arXiv:2605.16672v1 Announce Type: cross Abstract: Camera traps have become a common tool for wildlife monitoring efforts in ecological research and biodiversity conservation. Wildlife classification m

Multi-site PPG: An In-the-Wild Physiological Dataset from Emerging Multi-site Wearables

Model ReleasesDGX agent

arXiv:2605.17859v1 Announce Type: cross Abstract: Wearables are widely used for mobile health monitoring, and photoplethysmography (PPG) is a key sensing modality for heart rate and related physiologi

Offline Contextual Bandits in the Presence of New Actions

Model ReleasesDGX agent

arXiv:2605.18509v1 Announce Type: new Abstract: Automated decision-making algorithms drive applications such as recommendation systems and search engines. These algorithms often rely on off-policy con

Online Algorithms with Unreliable Guidance

TutorialsDGX agent

arXiv:2602.20706v2 Announce Type: replace Abstract: This paper introduces online algorithms with unreliable guidance (OAG), a model for ML-augmented online decision-making that cleanly separates the p

OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents

Model ReleasesDGX agent

arXiv:2506.16042v2 Announce Type: replace Abstract: Generative AI is being leveraged to solve a variety of computer-use tasks involving desktop applications. State-of-the-art systems have focused sole

Parallelizable memory recurrent units

ResearchDGX agent

arXiv:2601.09495v3 Announce Type: replace Abstract: With the emergence of massively parallel processing units, parallelization has become a desirable property for new sequence models. The ability to p

Patch-MoE Mamba: A Patch-Ordered Mixture-of-Experts State Space Architecture for Medical Image Segmentation

ResearchDGX agent

arXiv:2605.17719v1 Announce Type: new Abstract: CNN- and Transformer-based architectures have achieved strong performance in medical image segmentation, but CNNs are limited in modeling long-range dep

PEGRL: Improving Machine Translation by Post-Editing Guided Reinforcement Learning

Model ReleasesDGX agent

arXiv:2602.03352v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown strong promise for LLM-based machine translation, with recent methods such as GRPO demonstrating notable gains

Perfect Parallelization in Mini-Batch SGD with Classical Momentum Acceleration

Model ReleasesDGX agent

arXiv:2605.18609v1 Announce Type: new Abstract: Accelerating stochastic gradient methods with classical momentum schemes, such as Polyak's heavy ball, has proven highly successful in training large-sc

PFlow-T: A Persistence-Driven Forward Process for Topology-Controlled Generation

ResearchDGX agent

arXiv:2605.17555v1 Announce Type: cross Abstract: Current topology aware diffusion models face an architectural mismatch by using Gaussian noise for corruption while recovering structural features thr

PIMSM: Physics-Informed Multi-Scale Mamba for Stable Neural Representations under Distribution Shift

SafetyDGX agent

arXiv:2605.16351v1 Announce Type: cross Abstract: Scientific foundation models are expected to reuse representations under changes in dataset, acquisition protocol, and deployment domain, yet many seq

Plan First, Diffuse Later: Extrinsic Graph Guidance for Long-Horizon Diffusion Planning

Local AiDGX agent

arXiv:2605.16863v1 Announce Type: cross Abstract: Compositional diffusion models offer a promising route to long-horizon planning by denoising multiple overlapping sub-trajectories while ensuring that

Position: Weight Space Should Be a First-Class Generative AI Modality

ResearchDGX agent

arXiv:2605.18632v1 Announce Type: cross Abstract: Neural network checkpoints have quietly become a large-scale data resource: millions of trained weight vectors now exist, each encoding task-, domain-

PriHA: A RAG-Enhanced LLM Framework for Primary Healthcare Assistant in Hong Kong

Model ReleasesDGX agent

arXiv:2604.14215v2 Announce Type: replace-cross Abstract: To address the unsustainable rise in public health expenditures, the Hong Kong SAR Government is shifting its strategic focus to primary healt

Queue Length Regret Bounds for Contextual Queueing Bandits

Model ReleasesDGX agent

arXiv:2601.19300v2 Announce Type: replace Abstract: We introduce contextual queueing bandits, a new context-aware framework for scheduling while simultaneously learning unknown service rates. Individu

RaBiT: Residual-Aware Binarization Training for Accurate and Efficient LLMs

TutorialsDGX agent

arXiv:2602.05367v2 Announce Type: replace Abstract: Efficient deployment of large language models (LLMs) requires extreme quantization, forcing a critical trade-off between low-bit efficiency and perf

RadJEPA: Radiology Encoder for Chest X-Rays via Joint Embedding Predictive Architecture

TutorialsDGX agent

arXiv:2601.15891v2 Announce Type: replace Abstract: Recent advances in medical vision language models guide the learning of visual representations; however, this form of supervision is constrained by

REBAR: Reference Ethical Benchmark for Autonomy Readiness

Model ReleasesDGX agent

arXiv:2605.18423v1 Announce Type: new Abstract: As autonomous systems grow more advanced, objective metrics to evaluate their ethical and legal compliance are critical for informing end users of their

Responsible Federated LLMs via Safety Filtering and Constitutional AI

SafetyDGX agent

arXiv:2502.16691v2 Announce Type: replace Abstract: Recent research has increasingly focused on training large language models (LLMs) using federated learning, known as FedLLM. However, responsible AI

Rethinking GNNs and Missing Features: Challenges, Evaluation and a Robust Solution

Model ReleasesDGX agent

arXiv:2601.04855v2 Announce Type: replace-cross Abstract: Handling missing node features is a key challenge for deploying Graph Neural Networks (GNNs) in real-world domains such as healthcare and sens

Revisiting Long-term Time Series Forecasting: An Investigation on Linear Mapping

ApplicationsDGX agent

arXiv:2305.10721v2 Announce Type: replace-cross Abstract: Introduction: Long-term time series forecasting (LTSF) has gained significant attention in recent years. While various specialized designs exi

Ringmaster LMO: Asynchronous Linear Minimization Oracle Momentum Method

Model ReleasesDGX agent

arXiv:2605.18174v1 Announce Type: new Abstract: Muon has recently emerged as a strong alternative to AdamW for training neural networks, with encouraging large-scale pretraining results and growing ev

ROVR-Open-Dataset: A Large-Scale Depth Dataset for Autonomous Driving

Model ReleasesDGX agent

arXiv:2508.13977v3 Announce Type: replace Abstract: Depth estimation is a fundamental component of spatial perception for autonomous driving and other unmanned systems operating in open urban environm

S2Aligner: Pair-Efficient and Transferable Pre-Training for Sparse Text-Attributed Graphs

SafetyDGX agent

arXiv:2605.18579v1 Announce Type: new Abstract: Pre-training on text-attributed graphs (TAGs) is central to building transferable graph foundation models, where LLM-as-Aligner methods align graph and

SaaSBench: Exploring the Boundaries of Coding Agents in Long-Horizon Enterprise SaaS Engineering

Model ReleasesDGX agent

arXiv:2605.17526v1 Announce Type: cross Abstract: As autonomous coding agents become capable of handling increasingly long-horizon tasks, they have gradually demonstrated the potential to complete end

Safety Geometry Collapse in Multimodal LLMs and Adaptive Drift Correction

SafetyDGX agent

arXiv:2605.18104v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) often fail to transfer safety capabilities learned in the text modality to semantically equivalent non-text inp

Scalable Knowledge Editing for Mixture-of-Experts LLMs via Tensor-Structured Updates

Model ReleasesDGX agent

arXiv:2605.16686v1 Announce Type: new Abstract: Knowledge editing (KE) provides a lightweight alternative to repeated fine-tuning of LLMs. However, most existing KE methods target dense feed-forward l

Scale-Equivariant Generative Forecasting: Weight-Tied Dilated Convolutions, Wavelet Scattering Inputs, and Spectral-Consistency Training for Self-Similar Time Series

Model ReleasesDGX agent

arXiv:2605.17582v1 Announce Type: new Abstract: Many natural and engineered time series -- equity returns, climate anomalies, turbulent velocities, neural recordings, packet-level network traffic -- a

Scale-Invariant Neural Network Optimization: Norm Geometry and Heavy-Tailed Noise

ResearchDGX agent

arXiv:2605.18528v1 Announce Type: cross Abstract: A growing lesson from neural network optimization is that optimizer design should respect how the model is parametrized. Scale-invariant methods becom

SCAR: Self-Supervised Continuous Action Representation Learning

ResearchDGX agent

arXiv:2605.16412v1 Announce Type: cross Abstract: Despite the central role of action in embodied intelligence, learning transferable action representations from visual transitions remains a fundamenta

SEDD: Scalable and Efficient Dataset Deduplication with GPUs

Model ReleasesDGX agent

arXiv:2501.01046v4 Announce Type: replace Abstract: Dataset deduplication is widely recognized as a crucial preprocessing step that enhances data quality and improves the performance of large language

SEDualVLN: A Spatially-Enhanced Dual-System for Vision-Language Navigation

Local AiDGX agent

arXiv:2605.17249v1 Announce Type: new Abstract: Vision-Language Navigation (VLN) approaches have currently followed two primary paradigms: the end-to-end Vision-Language Model (VLM) policy fine-tuned

Seeking the Unfamiliar but Memorable: Conceptual Creativity as Meta-Learning

ApplicationsDGX agent

arXiv:2605.16477v1 Announce Type: cross Abstract: What does it mean to create a new concept, rather than retrieve a familiar one? Repeatedly sampling a generative model at the same prompt produces var

Self-Improving CAD Generation Agents with Finite Element Analysis as Feedback

Model ReleasesDGX agent

arXiv:2605.17448v1 Announce Type: cross Abstract: Computer-aided design (CAD) is the backbone of modern industrial design, yet learned CAD generators still fall short of real engineering pipelines: th

Setting the Stage: Text-Driven Scene-Consistent Image Generation

Model ReleasesDGX agent

arXiv:2512.12598v3 Announce Type: replace Abstract: We focus on the foundational task of Scene Staging: given a reference scene image and a text condition specifying an actor category to be generated

SIREM: Speech-Informed MRI Reconstruction with Learned Sampling

Model ReleasesDGX agent

arXiv:2605.18221v1 Announce Type: cross Abstract: Real-time magnetic resonance imaging (rtMRI) of speech production enables non-invasive visualization of dynamic vocal-tract motion and is valuable for

Skills on the Fly: Test-Time Adaptive Skill Synthesis for LLM Agents

Model ReleasesDGX agent

arXiv:2605.16986v1 Announce Type: cross Abstract: LLM agents benefit from reusable skills, yet test-time tasks often require guidance more specific than a static skill library can provide. We propose

SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution

Model ReleasesDGX agent

arXiv:2605.18401v1 Announce Type: cross Abstract: Long-horizon LLM agents leave traces that could become reusable experience, but raw trajectories are noisy and hard to govern. We treat Agent Skills a

SLEIGHT-Bench: A Benchmark of Evasion Attacks Against Agent Monitors

Model ReleasesDGX agent

arXiv:2605.16626v1 Announce Type: cross Abstract: Since autonomous coding agents generate complex behaviors at high-volume, we may want to use other LLMs to monitor actions to reduce the risk from dan

Small-scale photonic Kolmogorov-Arnold networks using standard telecom nonlinear modules

Model ReleasesDGX agent

arXiv:2604.08432v2 Announce Type: replace-cross Abstract: Photonic neural networks promise ultrafast inference, yet most architectures rely on linear optical meshes with electronic nonlinearities, rei

SomaliWeb v1: A Quality-Filtered Somali Web Corpus with a Matched Tokenizer and a Public Language-Identification Benchmark

Model ReleasesDGX agent

arXiv:2605.18232v1 Announce Type: cross Abstract: Somali is a Cushitic language of the Horn of Africa with ~25 million speakers, yet no documented dedicated Somali pretraining corpus with a companion

Sparse Training of Neural Networks based on Multilevel Mirror Descent

Model ReleasesDGX agent

arXiv:2602.03535v2 Announce Type: replace Abstract: We introduce a dynamic sparse training algorithm based on linearized Bregman iterations / mirror descent that exploits the naturally incurred sparsi

SurgLQA: Scalable Long-Horizon Surgical Video Question Answering

Model ReleasesDGX agent

arXiv:2605.17915v1 Announce Type: new Abstract: Surgical Video Question Answering (VideoQA) provides a promising paradigm for dynamic intraoperative interpretation, enabling real-time decision support

← Previous
1…555556557558559…1061
Next →