AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
30 Jul 2026

ChineseBERT: Chinese Pretraining Enhanced by Glyph and Pinyin Information

ResearchDGX agent

arXiv:2106.16038v3 Announce Type: replace Abstract: Recent pretraining models in Chinese neglect two important aspects specific to the Chinese language: glyph and pinyin, which carry significant synta

Cognitive Convergence: Deep Similarities Between Large Language Models and Human Cognition

ResearchDGX agent

arXiv:2607.26179v1 Announce Type: cross Abstract: LLMs are widely regarded as alien intelligences, systems whose cognitive operations are fundamentally unlike our own. Apparent similarities to human c

GEqTrain: A Configuration-Driven Framework for Retargeting Equivariant Graph Neural Networks Across 3D Scientific Tasks

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.19083v2 Announce Type: replace Abstract: Equivariant graph neural networks provide a powerful modeling language for three-dimensional scientific data, but their reuse is often limited by im

How Kimi K3 Engineered Its Way to the Frontier [R]

Model ReleasesDGX agent

Kimi K3 by Moonshot reached the frontier as an open-weight model. Artificial Analysis ranks it fourth of 580 models, behind only Claude Opus 5, Fable 5, and GPT-5.6 Sol. Moonshot released more than th

Nanbeige4.2-3B: I'm not impressed

Model ReleasesDGX agent

I've tested Nanbeige-4.2-3B. On paper, the benchmarks promise it blows away Qwen3.5-9B and Gemma4-12B. My goal was to have something very light and fast to replace Qwen3.6-35B (or finetunes thereof) f

Origins and mitigation of double descent in reduced order modeling

Local AiDGX agent

arXiv:2607.26414v1 Announce Type: cross Abstract: Latent low-dimensional structure in datasets of natural and engineered systems enables their sparse sensing, or full-state reconstruction from histori

Post-Training at the Edge of Detectability: A Game-Theoretic Approach to Fine-Tuning

Model ReleasesDGX agent

arXiv:2607.26358v1 Announce Type: new Abstract: Reinforcement learning (RL) fine-tuning is widely used in language model training to improve model performance on a target task while limiting drift fro

Prosody-driven Jailbreaks in Audio LLMs: A Controlled Study and Mechanistic Analysis

Model ReleasesDGX agent

arXiv:2607.26541v1 Announce Type: cross Abstract: Audio-capable foundation models enable end-to-end spoken interaction, but they also introduce safety risks beyond transcript content. It remains uncle

SERPO: Self-Evolving Rubric Policy Optimization for Open-Ended Test-Time Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.26873v1 Announce Type: new Abstract: Test-time reinforcement learning (TTRL) enables language models to self-evolve at inference time without labeled feedback. Existing methods rely on answ

29 Jul 2026

Cinematic Compositing Using Character-Environment-Harmonized Video Generation Models

ResearchDGX agent

arXiv:2606.20233v2 Announce Type: replace Abstract: Cinematic compositing aims to integrate green-screen characters into novel environments while maintaining physical and photometric realism. Previous

CycleVLA: Proactive Self-Correcting Vision-Language-Action Models via Subtask Backtracking and Minimum Bayes Risk Decoding

ResearchDGX agent

arXiv:2601.02295v2 Announce Type: replace Abstract: Current work on robot failure detection and correction typically operates in a post hoc manner, analyzing errors and applying corrections only after

DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space

Model ReleasesDGX agent

arXiv:2607.25675v1 Announce Type: new Abstract: Text-space optimization adapts large language models (LLMs) by editing external natural-language artifacts rather than model weights, so the optimized a

Evaluating Multi-Turn Multimodal Diagnostic Reasoning on Challenging Real-World Clinical Cases

Model ReleasesDGX agent

arXiv:2607.25933v1 Announce Type: cross Abstract: Clinical diagnostic evaluation should not only assess whether models can provide correct diagnoses, but also reflect the realities of clinical practic

Explanation-Bound Tool Execution for AI Agents: Server-Verified Action Claims Without Trusting Model Rationales

SafetyDGX agent

arXiv:2607.25364v1 Announce Type: new Abstract: Tool-using agents expose structured calls but commonly attach free-form rationales. Such rationales are neither authorization nor reliable introspection

Fast, accurate, and differentiable: a neural-network surrogate for NRSur7dq4 precessing binary black hole waveforms

Model ReleasesDGX agent

arXiv:2607.24960v1 Announce Type: cross Abstract: We present a neural network surrogate model that emulates the NRSur7dq4 gravitational waveform model for precessing binary black hole mergers. The sur

Finding Optimal Cost-Bounded Plan Reductions: Refined Model

ResearchDGX agent

arXiv:2607.25484v1 Announce Type: new Abstract: In some real applications a plan may later become unfeasible due to newly imposed budget constraints, yet, at the same time, using only the original act

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels

Model ReleasesDGX agent

arXiv:2607.24762v1 Announce Type: new Abstract: Machine learning models are increasingly embedded in everyday software, and most of their runtime is spent in a small set of compute kernels such as mat

Large Language Model for Operations Research Formulation Selection in Multi-Warehouse Inventory Allocation

SafetyDGX agent

arXiv:2607.25956v1 Announce Type: new Abstract: Multi-warehouse inventory allocation is typically formulated as a mixed-integer programming (MIP) problem, yet no single formulation consistently matche

Long-Term PM2.5 Forecasting Using a DTW-Enhanced CNN-GRU Model

ResearchDGX agent

arXiv:2510.22863v2 Announce Type: replace-cross Abstract: Reliable long-term forecasting of PM2.5 concentrations is critical for public health early-warning systems, yet existing deep learning approac

Measuring the State of Open Science in Transportation Using Large Language Models

ResearchDGX agent

arXiv:2601.14429v2 Announce Type: replace-cross Abstract: Open science initiatives have strengthened scientific integrity and accelerated research progress across many fields, but the state of their p

RIDGE: An Autonomous Framework for Validation and Method Discovery in LLM-Generated Option Pricing

Model ReleasesDGX agent

arXiv:2607.25199v1 Announce Type: cross Abstract: Automated code generation is becoming an important tool in quantitative finance, where large language models can generate option pricing implementatio

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement

Model ReleasesDGX agent

arXiv:2607.25886v1 Announce Type: cross Abstract: Recursive self-improvement requires turning evidence of model failures into better models. Data-centric post-training research entails diagnosing capa

Simulation-based parameter estimation via a combination of embedded normalizing flows and implied empirical probabilities under moment restrictions

Model ReleasesDGX agent

arXiv:2607.25026v1 Announce Type: cross Abstract: In this work, we present a simulation-based parameter estimation framework for a model defined by a computational simulation of a physical system. We

The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the Path Toward Fair Play

SafetyDGX agent

arXiv:2607.25425v1 Announce Type: new Abstract: Capture the Flag (CTF) competitions are among cybersecurity's most effective training grounds, developing practical skill across cryptography, web explo

Transformer Transformer: A Unified Model for Motion-Conditioned Robot Co-design

ResearchDGX agent

arXiv:2607.25798v1 Announce Type: new Abstract: An often overlooked factor of robot manipulation performance is the embodiment of the robot itself. Motivated by this problem, we study motion-condition

28 Jul 2026

A Coulomb Particle Model for Learning Kernel Attention in Transformers

SafetyDGX agent

arXiv:2607.23869v1 Announce Type: cross Abstract: Randomized features provide a scalable approximation to kernel machines, but their performance depends strongly on the choice of feature distribution.

A Few Words Go a Long Way: Language Guided Robot Policy Synthesis

Model ReleasesDGX agent

arXiv:2607.23784v1 Announce Type: cross Abstract: While vision-language-action models have demonstrated impressive zero-shot manipulation capabilities, they remain fundamentally black box policies tha

Agentic Reward Modeling: Verifying GUI Agent via Progressive Trajectory-Grounded Interaction

AgentsDGX agent

arXiv:2602.00575v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) provides a promising pathway for continuously advancing GUI agents, yet existing reward modeli

CALMRec: Causally Aligned Language Memory for Long-Horizon Recommendation

Model ReleasesDGX agent

arXiv:2607.23647v1 Announce Type: cross Abstract: Large language models (LLMs) can summarize heterogeneous user evidence in natural language, but current LLM recommenders often collapse enduring prefe

cMoLLM at Scale: Horizontal Scaling Laws for Mixture-of-LLMs

Model ReleasesDGX agent

arXiv:2607.22577v1 Announce Type: new Abstract: Scaling large language models (LLMs) has driven their success, yet dense Transformers couple capacity and computation: every parameter is activated for

Consistent Evidence, Robust Recognition: Faithful Attribution Regularization under Geometric Transformations

Model ReleasesDGX agent

arXiv:2607.23835v1 Announce Type: new Abstract: Attribution methods are widely used to characterize the evidence underlying model predictions, yet their potential to improve model behavior remains und

DeepSeek V4 Flash, up to 32 tok/s on AMD Ryzen AI MAX+ 395

Model ReleasesDGX agent

Hey fellow llamas. we have something new for Strix Halo owners we thought would be useful to share. i'll keep it short: We were able to fit DeepSeek V4 Flash plus its speculative draft on a single Ryz

DRC-Aid: Design-Rule Correction via Agentic Framework utilizing Inference-Time Large Language Models

Local AiDGX agent

arXiv:2607.22761v1 Announce Type: cross Abstract: Resolving Design Rule Violations (DRVs) in layouts entails an iterative loop of geometric edits and verification. We present DRC-Aid, a closed-loop ag

Hallucination Rates in Language Generation

Model ReleasesDGX agent

arXiv:2607.23361v1 Announce Type: cross Abstract: Language generation in the limit is an elegant model introduced by Kleinberg and Mullainathan [KM24] to formally study language generation by an algor

Intuitionistic j-Do-Calculus in Topos Causal Models

Local AiDGX agent

arXiv:2510.17944v2 Announce Type: replace-cross Abstract: In this paper, we generalize Pearl's do-calculus to an Intuitionistic setting called j-stable causal inference inside a topos of sheaves. Our

Kimi K3: Open Frontier Intelligence

Model ReleasesDGX agent

arXiv:2607.24653v1 Announce Type: new Abstract: We introduce Kimi K3, a 2.8T parameter Mixture-of-Experts model with 104 billion activated parameters, native vision capabilities, and a 1-million-token

LEX-EC: A Lexical Evidence-Channel Audit Framework for Zero-Shot LLM Personality Classification in Black-Box Settings

ResearchDGX agent

arXiv:2607.24435v1 Announce Type: cross Abstract: Large language models may easily assign personality labels from text, but model interpretability remains an open problem. To address this gap, we intr

MedFailBench: A Clinician-Built Open-Source Benchmark for Medical AI Safety Boundary Inspection

Model ReleasesDGX agent

arXiv:2607.15166v2 Announce Type: replace Abstract: Most medical AI benchmarks measure whether a model knows the correct answer. MedFailBench asks a different question: which safety boundary failed? W

MMOE: Modernizing Diffusion Transformers with Efficient Expert Design

Model ReleasesDGX agent

arXiv:2607.24665v1 Announce Type: new Abstract: Modern large language models scale successfully by pairing capacity growth with efficiency, keeping per-token and deployment costs under control as capa

Model Predictive Planner for UAV Navigation in Non-Convex Air Corridors

ResearchDGX agent

arXiv:2607.24369v1 Announce Type: new Abstract: This work presents a motion planning framework for UAV navigation in non-convex urban air corridors. The planner is based on a mixed-integer tracking mo

Open Your Model's Eyes: Video and Context-Aware Multimodal Backchannel Prediction

SafetyDGX agent

arXiv:2607.22729v1 Announce Type: cross Abstract: Backchannels, which signal listener states like empathy and understanding, are fundamental to natural human interaction. However, current approaches r

Pose-Aware Modeling to Mitigate Pose-Related Artifacts in Tactile Gloves

ResearchDGX agent

arXiv:2607.22964v1 Announce Type: new Abstract: Tactile gloves digitize contact and force during hand-object interactions, enabling robotics applications in dexterous manipulation, teleoperation, and

Similarity All The Way Up: Multilingual Generalization in LLMs Relies on Language-Level Similarity Structures

Model ReleasesDGX agent

arXiv:2607.22699v1 Announce Type: new Abstract: As Large Language Models (LLMs) grow more capable across diverse tasks, their (in)ability to generalize remains difficult to quantify and poorly underst

Spectral Dynamics of Semantic Drift in Clinical Multi-Agent Language Model Networks

SafetyDGX agent

arXiv:2607.22758v1 Announce Type: cross Abstract: The integration of iterative LLMs within multi-agent diagnostic frameworks requires a rigorous quantitative reevaluation of underlying communication t

To Erase, or Not to Erase: Robust Training-Free Concept Erasure with Preservation aware Adaptive Ranked Subspace Expansion

ResearchDGX agent

arXiv:2607.23492v1 Announce Type: new Abstract: Concept erasure techniques (CETs) edit text-to-image diffusion models to erase undesired targets such as NSFW content or copyrighted styles, while prese

Tokengeist: Multi-Turn Attribution Tracing in Agentic Conversations

Model ReleasesDGX agent

arXiv:2607.22610v1 Announce Type: new Abstract: When a language model produces a response in a multi-turn conversation, which tokens from prior turns shaped that answer, and how did those dependencies

Towards simultaneous decoding of kinetic and kinematic movement parameters during grasp and lift task by noninvasive brain imaging

Model ReleasesDGX agent

arXiv:2607.24081v1 Announce Type: cross Abstract: Brain-machine interfaces (BMIs) can assist individuals with limited mobility, such as stroke survivors or amputees. One of the key challenges in devel

Trustworthy Medical Segmentation: Uncertainty-Aware U-Net Evaluation Under Clinical Image Degradation

Model ReleasesDGX agent

arXiv:2607.22727v1 Announce Type: new Abstract: Medical image segmentation models often report high benchmark accuracy under ideal imaging conditions, yet their failures under clinical degradation can

What 'task oriented' models are folks running on N100 MiniPCs with 16GB of RAM and no GPU?

Local AiDGX agent

By 'task oriented', I dont really mean agentic, I mean no deep coding ability, no need for conversation. More things like classification, identification, simple interaction with web apps and APIs, etc

27 Jul 2026

b10148

Model ReleasesDGX agent

common: fix explicit -md precedence over draft sidecar resolution (#26165) common: fix explicit -md precedence over draft sidecar resolution Follow-up of #25955, an explicit --model-draft file given w

Bounding the Causal Impact of ML-assisted Decision-Making via Counterfactual Correctness

Model ReleasesDGX agent

arXiv:2607.21806v1 Announce Type: new Abstract: Predictive machine learning (ML) models are increasingly used to aid human decision-makers across various high-risk domains such as healthcare and crimi

DCS: A Unified Conditional Sensitivity Framework for Cross-Modal Copyright Infringement Detection

Local AiDGX agent

arXiv:2607.22035v1 Announce Type: new Abstract: Currently, most foundation models can reproduce or strongly depend on copyrighted training content, but output similarity alone is insufficient for infr

Enigma raises $71M to develop foundation models for robots

IndustryDGX agent

Engima Ltd., a provider of artificial intelligence software for robots, launched today with 71 million in funding. Index Ventures and Ribbit Capital jointly led the seed round with participation from

LatentFlow: Visual Analytics for Latent Space Analysis in Molecular Graph Neural Networks

ResearchDGX agent

arXiv:2607.21941v1 Announce Type: new Abstract: Chemists and materials scientists increasingly use machine learning models, such as graph neural networks (GNNs), to predict properties of molecules and

Layer-wise LoRA fine-tuning: a similarity metric approach

Model ReleasesDGX agent

arXiv:2602.05988v2 Announce Type: replace Abstract: Pre-training Large Language Models (LLMs) on web-scale datasets becomes fundamental for advancing general-purpose AI. In contrast, enhancing their p

MedKGent: A Large Language Model Agent Framework for Constructing Temporally Evolving Medical Knowledge Graph

AgentsDGX agent

arXiv:2508.12393v3 Announce Type: replace Abstract: The rapid expansion of medical literature challenges the scalable structuring of domain knowledge. Knowledge Graphs (KGs) offer a solution, yet curr

Nexus connects to the agentic harnesses your teams already use, whether that’s Claude Code, Codex, OpenCode, or your own custom tooling. It …

Model ReleasesDGX agent

Nexus connects to the agentic harnesses your teams already use, whether that’s Claude Code, Codex, OpenCode, or your own custom tooling. It gives you: → Intelligent routing, automatically matching eac

Unexpected use of local llm

Local AiDGX agent

I was refreshing my youtube and found out my favourite reviewer uploaded a battery test of 78 smartphones: https://youtu.be/MpgUFrsIWSQ the author said they started using robotic arm to simulate a per

We could really use Qwen3.8 in 27B, 35B, 122B and 397B sizes

Model ReleasesDGX agent

Instead of 2T+ models, continuing to release highly capable small to medium size LLMs would really help to keep this community vibrant. Hardly anyone can even dream of running the recent 1.5-2T+ beast

You can now fine-tune my 3.96M-parameter TTS on your own voice or language

Model ReleasesDGX agent

When I released Inflect v2 last week, I thought most people would ask whether a TTS model this small actually sounded decent. Instead, I kept getting two questions: “Can I train it on my own voice?” “

← Previous
1…246247248249250…1018
Next →