AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,429Total entries
1Added by human
88,428Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,648 results
24 Jun 2026

Dimensionality Reduction of QAOA Parameter Space with Kernel PCA for Max-Cut

Model ReleasesDGX agent

arXiv:2606.23718v1 Announce Type: cross Abstract: The Quantum Approximate Optimization Algorithm (QAOA) is a leading variational algorithm for combinatorial optimization on near term quantum devices.

Dirac-Frenkel dynamics with inertia for nonlinearly parametrized solutions of evolution problems

Model ReleasesDGX agent

arXiv:2606.24769v1 Announce Type: cross Abstract: Even when Dirac-Frenkel dynamics determine a well-defined evolution in function space, the corresponding parameter dynamics can be non-unique or ill-c

DLTPose: 6DoF Pose Estimation From Accurate Dense Surface Point Estimates

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2504.07335v3 Announce Type: replace Abstract: We propose DLTPose, a novel method for 6DoF object pose estimation from RGBD images that combines the accuracy of sparse keypoint methods with the r

EchoFoley: Event-Centric Hierarchical Control for Video Grounded Creative Sound Generation

Model ReleasesDGX agent

arXiv:2512.24731v2 Announce Type: replace Abstract: Sound effects build an essential layer of multimodal storytelling, shaping the emotional atmosphere and the narrative semantics of videos. Despite r

EComAgentBench: Benchmarking Shopping Agents on Long-Horizon Tasks with Distributed Hidden Intent

Model ReleasesDGX agent

arXiv:2606.17698v2 Announce Type: replace Abstract: As LLM-based shopping agents enter production, existing benchmarks fail to capture how a shopper's requirements arrive: stated implicitly in the que

Faithful by Construction: Claim-Anchored Attribution for Multi-Document Summarization

Local AiDGX agent

arXiv:2606.23989v1 Announce Type: cross Abstract: End-to-end large language models (LLMs) produce fluent multi-document summaries but remain prone to hallucination, and the attributions they offer are

FlowerDance: MeanFlow for Efficient and Refined 3D Dance Generation

Model ReleasesDGX agent

arXiv:2511.21029v3 Announce Type: replace Abstract: Music-to-dance generation aims to translate auditory signals into expressive human motion, with broad applications in virtual reality, choreography,

GUI vs. CLI: Execution Bottlenecks in Screen-Only and Skill-Mediated Computer-Use Agents

Model ReleasesDGX agent

arXiv:2606.24551v1 Announce Type: new Abstract: Computer-use agents can execute software tasks through either graphical interfaces or programmatic command interfaces, but existing evaluations confound

HelloTwin launches ‘Digital Authority’ to bring governed AI agents to the enterprise

Model ReleasesDGX agent

HelloTwin.ai GmbH today announced what it calls an accountable artificial intelligence AI twin that holds business intelligence and goals in a single source of truth. HelloTwin said it built its data

Inclusive Interactive Collisions for Multi-View Consistent Compositional 3D Generation

TutorialsDGX agent

arXiv:2606.24206v1 Announce Type: cross Abstract: Recent breakthroughs in 3D generation have advanced notably with the development of text-to-image diffusion model. However, existing methods remain tw

Invariant Graph Representations for Continuous-Time Dynamic Graphs Under Distribution Shifts

TutorialsDGX agent

arXiv:2405.19062v2 Announce Type: replace-cross Abstract: Continuous-Time Dynamic Graphs (CTDGs) enable fine-grained modeling of evolving relational systems. However, most existing CTDG representation

Layer-wise Probing of wav2vec 2.0 and Whisper for Consonant Cluster Reduction in African American English

ResearchDGX agent

arXiv:2606.23948v1 Announce Type: new Abstract: Self-supervised and supervised speech models are increasingly used to investigate which linguistic information their internal representations encode, an

Learning the Koopman Operator using Attention Free Transformers

Model ReleasesDGX agent

arXiv:2606.23957v1 Announce Type: new Abstract: Learning Koopman operators with autoencoders enables linear prediction in a latent space, but long-horizon rollouts often drift off the learned manifold

LoT-Pass: Long-term-robust Image Watermarking for Image to Video Generation

Model ReleasesDGX agent

arXiv:2509.17773v2 Announce Type: replace Abstract: The rapid progress of image-guided video generation (I2V) has raised concerns about its potential misuse in misinformation and fraud, underscoring t

MuTRAP: Multi-trigger Trojans Attacking Robot Task Planning Systems

ResearchDGX agent

arXiv:2504.17070v3 Announce Type: replace-cross Abstract: Robots need task planning methods to achieve goals that require more than one action. Recently, large pretrained models have demonstrated impr

NatureBench: Can Coding Agents Match the Published SOTA of Nature-Family Papers?

Model ReleasesDGX agent

arXiv:2606.24530v1 Announce Type: new Abstract: We introduce NatureBench, a cross-discipline benchmark of 90 tasks distilled from peer-reviewed Nature-family publications, designed to evaluate whether

Navigating User Behavior toward Personalized Multimodal Generation

Model ReleasesDGX agent

arXiv:2606.24196v1 Announce Type: new Abstract: Modern AIGC pipelines deliver high-fidelity images and videos but presuppose a well-formed creation instruction, while end users rarely articulate visua

PEARL: Self-Evolving Assistant for Time Management with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.11957v4 Announce Type: replace Abstract: Overlapping calendar invitations force busy professionals to repeatedly decide which meetings to attend, reschedule, or decline. We refer to this pr

PROTECT-90: A Fault Dataset for Power System Protection

Model ReleasesDGX agent

arXiv:2606.24298v1 Announce Type: cross Abstract: The increasing interest in data-driven methods for power system protection is accompanied by a lack of standardized, publicly available high-voltage w

Quantifying Prior Dominance in RAG Systems

ResearchDGX agent

arXiv:2606.23695v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) grounds Large Language Models in external knowledge, yet current evaluations rely on discrete heuristics that suf

QuechuaTok: Morphological Boundary Accuracy as a Necessary Metric for Tokenizer Evaluation in Agglutinative Low-Resource Languages

Model ReleasesDGX agent

arXiv:2606.23943v1 Announce Type: new Abstract: Tokenization is a foundational step in NLP pipelines, yet standard evaluation metrics such as fertility rate fail to capture morphological correctness f

Random Rule Forest (RRF): Interpretable and Manageable Ensembles of LLM-Generated Questions for Predicting Success from Unstructured Data

Model ReleasesDGX agent

arXiv:2505.24622v3 Announce Type: replace Abstract: Many high-stakes screening tasks require predicting rare outcomes from unstructured text, where errors are costly and decisions must be auditable. W

RAVEN: A Regime-Aware Variable-context Expert Network for Financial Time Series Forecasting

Model ReleasesDGX agent

arXiv:2606.24062v1 Announce Type: cross Abstract: Financial time series forecasting presents structural challenges absent from standard benchmarks. Log-returns are non-stationary, exhibit exceptionall

RE4: Transformation-aware Imitation of Object Interactions Using Manipulation Modes

Model ReleasesDGX agent

arXiv:2606.24403v1 Announce Type: cross Abstract: Object interaction tasks have been a focus of advances in imitation learning. End-to-end methods, dominated by diffusion and flow-based variants have

Removing Noise, not Finding Gold: Quality Filtering for Large-Scale Pretraining

ResearchDGX agent

arXiv:2510.00866v3 Announce Type: replace-cross Abstract: Large-scale models are pretrained on massive web-crawled datasets containing documents of mixed quality, making data filtering essential. A po

Right alongside Cursor, Devin Desktop (Windsurf) and CLI now support GLM-5.2 as well. FrontierCode Extended is a benchmark we care deeply ab…

Model ReleasesDGX agent

Right alongside Cursor, Devin Desktop (Windsurf) and CLI now support GLM-5.2 as well. FrontierCode Extended is a benchmark we care deeply about for real-world engineering tasks, so it's great to see G

RoPE-Aware Bit Allocation for KV-Cache Quantization

Model ReleasesDGX agent

arXiv:2606.24033v1 Announce Type: cross Abstract: Existing low-bit KV-cache quantizers often treat each cached key as a flat vector. Under RoPE, however, a key's contribution to a future attention log

SAFARI: Scaling Long Horizon Agentic Fault Attribution via Active Investigation

Model ReleasesDGX agent

arXiv:2606.24626v1 Announce Type: new Abstract: As autonomous agents tackle increasingly complex multi-step, multi-agent tasks, their execution trajectories have scaled beyond the constraints of even

SEAL: Searching Expandable Architectures for Incremental Learning

SafetyDGX agent

arXiv:2505.10457v3 Announce Type: replace-cross Abstract: Incremental learning is a machine learning paradigm where a model learns from a sequential stream of tasks. This setting poses a key challenge

SER: Learning to Ground Video Reasoning with Semantic Evidence Rewards

Model ReleasesDGX agent

arXiv:2606.24726v1 Announce Type: new Abstract: Video MLLMs often struggle with fine-grained spatio-temporal reasoning, sometimes generating correct answers based on irrelevant frames or objects. Alth

simonw/browser-compat-db

Model ReleasesDGX agent

simonw/browser-compat-db Inspired by Mozilla's new MDN MCP service - source code here - I decided to try converting their comprehensive mdn/browser-compat-data repository full of browser compatibility

Societal Alignment Frameworks Can Improve LLM Alignment

SafetyDGX agent

arXiv:2503.00069v2 Announce Type: replace-cross Abstract: Recent progress in large language models (LLMs) has focused on producing responses that meet human expectations and align with shared values -

Stochastic Expectation Maximization for Robust State-Space Radio Interferometric Imaging

ResearchDGX agent

arXiv:2606.23944v1 Announce Type: cross Abstract: State--space models provide a flexible framework for analyzing dynamical systems, yet they often rely on Gaussian assumptions that fail to capture hea

SURGELLM: Rethinking Multi-Task Evaluation through Task-Aware Feature Gating with Class-Balanced Normalization

Model ReleasesDGX agent

arXiv:2606.24259v1 Announce Type: cross Abstract: Fine-tuned encoders deployed across heterogeneous NLP tasks face three compounding problems: mismatched inductive biases, class-imbalance corruption o

Synergizing Physically Constrained MCMC and Chemical-Informed Gaussian Processes for Reaction Network Discovery

Model ReleasesDGX agent

arXiv:2606.23757v1 Announce Type: cross Abstract: Extracting interpretable governing equations from sparse, noisy chemical time-series data remains difficult because discrete reaction topology and con

The Latent Bridge: A Continuous Slow-Fast Channel for Real-Time Game Agents

AgentsDGX agent

arXiv:2606.24470v1 Announce Type: new Abstract: A real-time agent for general computer use - with games as the most demanding case - must act within tens of milliseconds while still planning over seco

TIGER: Taming Identity, Geometry, and Generative Priors for High-Quality Face Video Restoration

Model ReleasesDGX agent

arXiv:2606.24336v1 Announce Type: new Abstract: Face Video Restoration (FVR) aims to recover high-fidelity facial videos from degraded input while preserving identity and semantic consistency across f

Token-to-Token Alignment of Text Embeddings for Semantic Blending

SafetyDGX agent

arXiv:2606.24021v1 Announce Type: new Abstract: In modern generative models, images are specified and controlled through text prompts. In practice, images are generated from sequences of tokens derive

Towards Version-aware Operations and Transaction Memories for Multi-layer MeMo

ResearchDGX agent

arXiv:2606.24040v1 Announce Type: cross Abstract: MeMo proposes language models with explicit multi-layer correlation matrix memories (CMMs), where memorization, retrieval, and forgetting are architec

Try Kimi K2.7 and GLM 5.2 for free in Devin Desktop and CLI

Model ReleasesDGX agent

Try Kimi K2.7 and GLM 5.2 for free in Devin Desktop and CLI Kimi K2.7 Code and GLM 5.2 are available in Devin Desktop and CLI Both perform strongly on FrontierCode Extended, our benchmark for real-wor

Tuning without Peeking: Provable Generalization Bounds and Robust LLM Post-Training

Model ReleasesDGX agent

arXiv:2507.01752v4 Announce Type: replace-cross Abstract: Gradient-based optimization is the workhorse of deep learning, offering efficient and scalable training via backpropagation. However, exposing

VistaRef: Boosting Visual Spatial Orientation Awareness for Pointing-to-Object Detection

Local AiDGX agent

arXiv:2606.24498v1 Announce Type: new Abstract: Grounding deictic gestures in natural images is fundamental to AR and human-robot collaboration, providing a basis for seamless spatial interaction. Whi

23 Jun 2026

A Generalized Formalism of Auto-Regressive Decoding for Speech Processing

ResearchDGX agent

arXiv:2606.20714v1 Announce Type: cross Abstract: In speech processing, most state-of-the-art sequence prediction models rely on auto-regressive (AR) strategies to generate output sequences based on t

A Hybrid, Multi-Layered Pipeline for Phishing and Threat Classification: Independently Validated URL and NLP Engines with a Calibrated Multi-Channel Fusion Stage

Model ReleasesDGX agent

arXiv:2606.21690v1 Announce Type: cross Abstract: Phishing is a multi-modal threat. We present a hybrid pipeline that scores each modality with its own engine and fuses the results. Three engines are

A Latent Representation Learning Framework for Hyperspectral Image Emulation in Remote Sensing

Model ReleasesDGX agent

arXiv:2603.21911v2 Announce Type: replace Abstract: Synthetic hyperspectral image (HSI) generation is essential for large-scale simulation, algorithm development, and mission design, yet traditional r

A Smart Classroom Behavior Analysis Framework with a New Highly Congested Classroom Dataset

Model ReleasesDGX agent

arXiv:2606.21568v1 Announce Type: new Abstract: Student behavior detection is important for intelligent classroom analysis but remains challenging in large-class scenarios due to dense instance co-occ

A Standard Processing Pipeline for High-accuracy Measurement of Few-shot Regression on Laser Induced Breakdown Spectroscopy

Model ReleasesDGX agent

arXiv:2606.21960v1 Announce Type: new Abstract: Laser-induced breakdown spectroscopy (LIBS) faces challenges in high-accuracy quantitative measurement under few-shot scenarios due to spectral noise an

A Stitch in Time Saves Nine: Preserving Policy Compatibility Under Perception Updates in End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2606.21509v1 Announce Type: new Abstract: End-to-end autonomous driving systems tightly couple perception and decision-making through latent representations. Consequently, updates to perception

An Analysis of Untrained Deep Reservoir Networks for Audio Surveillance

HardwareDGX agent

arXiv:2606.22218v1 Announce Type: cross Abstract: In this paper, we investigate untrained recurrent models from the Reservoir Computing (RC) paradigm for audio surveillance, focusing on bidirectional

An Efficient and Effective Architecture for Large-Scale Traffic Prediction via Geometry-Adaptive Square Partitioning

ApplicationsDGX agent

arXiv:2606.21072v1 Announce Type: new Abstract: Traffic prediction is a core task in intelligent transportation systems and urban-scale decision making. Despite the effectiveness of mainstream neural-

Anticipating the Optimism Gap: Predicting Distribution-Shift Degradation of RF-Impairment Detectors from In-Distribution Statistics

Model ReleasesDGX agent

arXiv:2606.22054v1 Announce Type: cross Abstract: Detectors for GNSS radio-frequency impairments (jamming, spoofing, multipath) are usually reported with a single AUC measured on the distribution they

Boundary-by-Mask: Few-Shot Instance Segmentation with Mask-Conditioned Boundary Learning for Texture-Poor Industrial Parts

ResearchDGX agent

arXiv:2606.21594v1 Announce Type: new Abstract: Recent advances in large pre-trained models have led to remarkable progress in instance segmentation on general images. However, industrial scenarios re

BranchShine: Compact Raw-Audio-to-IPA Transcription with a RoPE E-Branchformer Encoder

Model ReleasesDGX agent

arXiv:2606.22824v1 Announce Type: new Abstract: Speech-to-IPA transcription is useful when the desired output is pronunciation rather than orthographic text, but competitive multilingual systems are o

BYOK is now live in the GitHub Copilot App! Works with @ollama, foundry, and any OAI completions or Anthropic compatible messages endpoint. …

Local AiDGX agent

GitHub Copilot App now supports Bring Your Own Key (BYOK) functionality, allowing users to integrate local and third-party AI models including Ollama, Foundry, and any OpenAI-compatible or Anthropic-c

Category-Adaptive Cross-Modal Semantic Refinement and Transfer for Open-Vocabulary Multi-Label Recognition

ResearchDGX agent

arXiv:2412.06190v2 Announce Type: replace Abstract: Benefiting from the generalization capability of CLIP, recent vision language pre-training (VLP) models have demonstrated the ability to capture a w

CFPO: Counterfactual Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2606.23206v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated remarkable capabilities in multimodal reasoning. However, prevailing reinforcement learning (RL)

Cloak: Zero-Shot Cross-Embodiment Manipulation by Masking the End-Effector from the VLA

ResearchDGX agent

arXiv:2606.22836v1 Announce Type: new Abstract: We present Cloak, a training recipe that endows a Vision-Language-Action (VLA) model with zero-shot cross-embodiment transfer by cloaking the end-effect

CodePercept: Code-Grounded Visual STEM Perception for MLLMs

Model ReleasesDGX agent

arXiv:2603.10757v2 Announce Type: replace Abstract: When MLLMs fail at Science, Technology, Engineering, and Mathematics (STEM) visual reasoning, a fundamental question arises: is it due to perceptual

CoDMD: Copula-aware Distribution Matching Distillation for Fast Video Generation

Local AiDGX agent

arXiv:2606.21982v1 Announce Type: new Abstract: Few-step distillation for video diffusion models has attracted significant attention, driven by the urgent demand for efficient deployment in real-world

Coherence Under Commitment: Probing Generalization and Vacuous Memorization in LLM Logical Reasoning

ResearchDGX agent

arXiv:2606.21083v1 Announce Type: cross Abstract: Large language models (LLMs) deployed for logical reasoning in knowledge-intensive domains exhibit a subtle but critical failure: coherence can be vac

← Previous
1…531532533534535…1061
Next →