AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
All
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,620 results
Model Releases

Meta rolls out Plus plans for Instagram, Facebook, and WhatsApp globally, will test 7.99/mo. and 19.99/mo. Meta AI plans, a $49.99/mo. creator plan, and more (Sarah Perez/TechCrunch)

DGX agent

Sarah Perez / TechCrunch: Meta rolls out Plus plans for Instagram, Facebook, and WhatsApp globally, will test 7.99/mo. and 19.99/mo. Meta AI plans, a $49.99/mo. creator plan, and more — Meta is doubli

model-releasestechmeme
27 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

METATR: A Multilingual, Evolving Benchmark for Automatic Text Recognition

DGX agent

arXiv:2605.26712v1 Announce Type: new Abstract: Benchmarks that reflect the diversity and complexity of real-world documents are essential for accurately evaluating Automatic Text Recognition (ATR) sy

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

MobileExplorer: Accelerating On-Device Inference for Mobile GUI Agents via Online Exploration

DGX agent

arXiv:2605.26546v1 Announce Type: new Abstract: Mobile graphical user interface (GUI) agents enable AI models to autonomously operate smartphones on behalf of users. However, most existing systems foc

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

MobileMoE: Scaling On-Device Mixture of Experts

DGX agent

arXiv:2605.27358v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the de facto architecture for hundred-billion-parameter language models, yet its advantages at sub-billion scales

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Model discovery for dynamical systems with complex-valued product units

DGX agent

arXiv:2605.27158v1 Announce Type: new Abstract: Discovering the governing equations of a dynamical system from observed trajectories provides deeper insight into its structure than mere prediction of

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Model-Harness-Task fit! it’s clear that RL post-training produces a model-harness fit via tool shapes and prompting as models are trained wi…

DGX agent

Model-Harness-Task fit! it’s clear that RL post-training produces a model-harness fit via tool shapes and prompting as models are trained with the harness in the loop. Mentioned this in a previous Lan

model-releasesharrison-chase--x
27 May 2026
Model Releases

Modeling Dynamic Mixtures of Time-Delay Systems from Streaming Time Series

DGX agent

arXiv:2605.26191v1 Announce Type: cross Abstract: This research addresses the problem of adaptive modeling in time-series data streams with clear input-output relationships. This problem is challengin

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

MolPIF: A Parameter Interpolation Flow Model for Molecule Generation

DGX agent

arXiv:2507.13762v4 Announce Type: replace Abstract: Motivation: Structure-based drug design (SBDD) has advanced with deep generative models, but bridging the gap between continuous atomic coordinates

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

More Expressive Feedforward Layers: Part I. Token-Adaptive Mixing of Activations

DGX agent

arXiv:2605.26647v1 Announce Type: cross Abstract: Feedforward network (FFN) layers account for a large fraction of parameters and nonlinear expressivity in Transformer-based large language models (LLM

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

MRT: Masked Region Transformer for Layered Image Generation and Editing at Scale

DGX agent

arXiv:2605.27235v1 Announce Type: new Abstract: Layered image generation and editing is a fundamental capability that enables layer-wise reuse, editing, and composition of generated visual content, an

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

MTL-FNO: A Lightweight Multi-Task Fourier Neural Operator for Sparse Field Reconstruction

DGX agent

arXiv:2605.26718v1 Announce Type: new Abstract: Efficient onboard multi-field sparse reconstruction is essential for the autonomous operation of aerospace vehicles. While existing deep learning models

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Multi-Agent Causal Discovery Using Large Language Models

DGX agent

arXiv:2407.15073v4 Announce Type: replace Abstract: Causal discovery aims to identify causal relationships between variables and is a fundamental problem across the sciences. Traditional statistical c

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

MULTISEISMO: A Multimodal Seismic Dataset and Model for Cross-Modal Seismic Understanding

DGX agent

arXiv:2605.26320v1 Announce Type: cross Abstract: The application of generalist multimodal models (GMMs) to specialized scientific domains remains limited due to the scarcity of comprehensive domain-s

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Near-Optimal Regret in Adversarial Kernel Bandits

DGX agent

arXiv:2605.26585v1 Announce Type: new Abstract: We study the adversarial kernel bandit problem, in which the loss at each round is induced by an arbitrary bounded element of a reproducing kernel Hilbe

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Negligible in Size, Significant in Effect: On Scale Vectors in Large Language Models

DGX agent

arXiv:2605.26895v1 Announce Type: cross Abstract: Normalization layers in modern large language models (LLMs) consist of a deterministic normalization operation and a learnable scale vector. While the

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

NestedKV: Nested Memory Routing for Long-Context KV Cache Compression

DGX agent

arXiv:2605.26678v1 Announce Type: new Abstract: Long-context language models are limited by the memory footprint of the key-value (KV) cache. Existing training-free KV compression methods usually rank

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Neural Autoregressive Control Variates for the Quantum Monte Carlo Sign Problem

DGX agent

arXiv:2605.26814v1 Announce Type: cross Abstract: We train a pair of autoregressive models to construct zero-mean control variates to mitigate the sign problem in quantum Monte Carlo simulations. The

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Not All Disagreement Is Learnable: Token Teachability in On-Policy Distillation

DGX agent

arXiv:2605.26844v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own rollouts with token-level teacher supervision. Recent selective OPD methods exploit the non-uni

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

O-MARC: Omni Memory-Augmented Compression Distillation for Efficient Video Understanding

DGX agent

arXiv:2605.26584v1 Announce Type: new Abstract: Omnimodal large language models enable unified audio video understanding, but long joint token sequences make inference costly, and existing benchmarks

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning

DGX agent

arXiv:2505.17163v2 Announce Type: replace-cross Abstract: Recent advancements in multimodal slow-thinking systems have demonstrated remarkable performance across various visual reasoning tasks. Howeve

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection

DGX agent

arXiv:2508.01253v2 Announce Type: replace Abstract: Existing studies typically investigate domain shift and category shift as independent problems, however, in real-world scenarios, the two types of s

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Omanic: Towards Step-wise Evaluation of Multi-hop Reasoning in Large Language Models

DGX agent

arXiv:2603.16654v2 Announce Type: replace-cross Abstract: Evaluating the reasoning abilities of large language models (LLMs) solely from final answers can obscure failures in intermediate steps, espec

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

OMD-GraphRAG: Enhancing GraphRAG with Ontology-Guided Extraction, Multi-Dimensional Clustering and Dual-Channel Fusion

DGX agent

arXiv:2603.25152v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems face significant challenges in complex reasoning, multi-hop queries, and domain-specific QA. While exis

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

OmniInteract: Benchmarking Real-World Streaming Interaction for Real-Time Omnimodal Assistants

DGX agent

arXiv:2605.26485v1 Announce Type: cross Abstract: We introduce OmniInteract, a streaming benchmark for real-time omnimodal large language models evaluated through native online inference over audio-vi

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

OmniRetriever: Any-to-Any Audio-Video-Text Retrieval via Fusion-as-Teacher Distillation

DGX agent

arXiv:2605.26641v1 Announce Type: new Abstract: Unified multimodal embedding spaces have become the standard interface for cross-modal retrieval and multimodal RAG, and recent audio-video-text (AVT) e

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

OmniToM: Benchmarking Theory of Mind in LLMs via Explicit Belief Modeling

DGX agent

arXiv:2605.26322v1 Announce Type: new Abstract: Theory of Mind (ToM), the ability to infer others' knowledge, intentions, and emotions, is commonly evaluated in large language models (LLMs) using end-

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

On the Hidden Costs of Counterfactual Knowledge Training in LLM Unlearning

DGX agent

arXiv:2605.27083v1 Announce Type: new Abstract: Counterfactual tuning (CFT) has emerged as a promising paradigm for Large Language Model (LLM) unlearning by training models to generate alternative fic

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

On the Sensitivity of Instruction-tuned LLMs to Harmful Sentences in Long Inputs

DGX agent

arXiv:2510.05864v2 Announce Type: replace Abstract: Large language models (LLMs) increasingly operate on long inputs, yet their behavior when harmful sentences are sparsely embedded within such inputs

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

ORLoopBench: Solver-in-the-Loop Benchmarks for Self-Correction and Behavioral Rationality in Operations Research

DGX agent

arXiv:2601.21008v3 Announce Type: replace-cross Abstract: Operations Research practitioners debug infeasible models through an iterative process: inspecting Irreducible Infeasible Subsystems ( IIS), i

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

OSMa-Bench++: Toward Open-Ended Benchmarking of Semantic Mapping for Manipulation with Prompt-Generated Synthetic Scenes

DGX agent

arXiv:2605.26831v1 Announce Type: new Abstract: Semantic mapping methods are increasingly used as intermediate scene representations for downstream robotic reasoning and manipulation, yet their evalua

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

PashtoTTS-Bench: automated screening for low-resource non-Latin-script text-to-speech

DGX agent

arXiv:2605.26978v1 Announce Type: new Abstract: Text-to-speech (TTS) evaluation for low-resource non-Latin-script languages can fail when it relies on a single ASR round-trip word error rate (WER). A

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

PaTAS: A Framework for Trust Propagation in Neural Networks Using Subjective Logic

DGX agent

arXiv:2511.20586v4 Announce Type: replace Abstract: Trustworthiness has become a key requirement for the deployment of artificial intelligence systems in safety-critical applications. Conventional eva

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Periodic Topological Deep Learning for Polymer Design and Discovery

DGX agent

arXiv:2605.26833v1 Announce Type: cross Abstract: Polymers underpin applications across energy, healthcare, and materials science, yet their vast chemical space makes systematic discovery challenging.

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark

DGX agent

arXiv:2506.00250v4 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved remarkable performance on a wide range of Natural Language Processing (NLP) benchmarks, often surpassing

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

PersLitEval: Fine-grained Benchmark and Evaluation of LLMs on Persian Literature Questions

DGX agent

arXiv:2605.27015v1 Announce Type: new Abstract: Despite impressive multilingual capabilities, large language models (LLMs) remain poorly evaluated on literary knowledge in non-English languages. We in

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History

DGX agent

arXiv:2602.17003v2 Announce Type: replace-cross Abstract: Large language models have advanced web agents, yet current agents lack personalization capabilities. Since users rarely specify every detail

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

'PhyWorldBench': A Comprehensive Evaluation of Physical Realism in Text-to-Video Models

DGX agent

arXiv:2507.13428v3 Announce Type: replace-cross Abstract: Video generation models have achieved remarkable progress in creating high-quality, photorealistic content. However, their ability to accurate

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

PIDM-DP: Physics-Informed Diffusion with Dormand-Prince Integration for Chaotic System Identification and State Reconstruction across Multiple Dynamical Regimes

DGX agent

arXiv:2605.26619v1 Announce Type: new Abstract: Reconstructing continuous state trajectories of chaotic dynamical systems from sparse, noisy observations remains a fundamental open problem in nonlinea

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis

DGX agent

arXiv:2605.27258v1 Announce Type: cross Abstract: Building state-of-the-art text-to-speech (TTS) systems typically demands millions of hours of proprietary data and complex multi-stage architectures,

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

PitchBench: Measuring Pitch Hearing in Audio-Language Models

DGX agent

arXiv:2605.26176v1 Announce Type: cross Abstract: Audio-language models (ALMs) are increasingly used in real-world applications that require understanding music, from music tutoring and transcription

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Position: AI Safety Requires Effective Controllability

DGX agent

arXiv:2605.27117v1 Announce Type: new Abstract: AI safety is still largely framed as alignment: training models to follow human preferences, safety policies, and normative constraints. That framing ha

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

PRBench: A Standardized Probabilistic Robustness Benchmark

DGX agent

arXiv:2511.01724v3 Announce Type: replace Abstract: Deep learning models are notoriously vulnerable to imperceptible perturbations. Most existing research centers on adversarial robustness (AR), which

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Pretrained Approximators for Low-Thrust Trajectory Cost and Reachability

DGX agent

arXiv:2605.26790v1 Announce Type: new Abstract: Low-thrust trajectory design relies heavily on repeated evaluations of fuel consumption and transfer feasibility, which require expensive optimal contro

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

PRISM: A Multi-Dimensional Benchmark for Evaluating LLM Peer Reviewers

DGX agent

arXiv:2605.26730v1 Announce Type: new Abstract: The rapid growth in submissions to machine learning venues has strained the scientific peer-review system and intensified interest in LLM-based automate

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

PRISM: Position-encoded Regressive Inverse Spectral Model for Multilayer Thin-Film Design

DGX agent

arXiv:2605.26502v1 Announce Type: new Abstract: The inverse problem of multilayer thin-film optical coatings design represents a complex combinatorial-continuous optimization challenge. We present PRI

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Probing Cultural Awareness in LLMs: A Case Study of Cross-Culture Aesthetic Stylistics

DGX agent

arXiv:2605.27296v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in diverse cultural contexts, yet their ability to master aesthetic stylistics, i.e., the strateg

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Prompt Injection Detection is Regime-Dependent: A Deployment-Aware Evaluation with Interpretable Structural Signals

DGX agent

arXiv:2605.26999v1 Announce Type: new Abstract: Prompt injection poses a critical threat to the safe deployment of large language models, yet existing detection approaches are typically evaluated unde

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Provably Communication-Efficient and Privacy-Preserving Federated Graph Neural Networks

DGX agent

arXiv:2605.26243v1 Announce Type: new Abstract: Graph neural networks (GNNs) achieve strong performance on relational data, but real-world graphs are often distributed across organizations that cannot

model-releasesarxiv-cs-lg
27 May 2026
← Previous
1…258259260261262…472
Next →