AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,555 results
Model Releases

SenWorld: A Digital-Twin Simulation for Generating Context-Rich Evaluation Data

DGX agent

arXiv:2607.19949v1 Announce Type: new Abstract: Smartphone personal assistants reason over longitudinal personal data, yet evaluating them requires context-rich evaluation data whose correct answers a

model-releasesarxiv-cs-ai
23 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Simultaneous Speech-to-Speech Translation Without Aligned Data

DGX agent

arXiv:2602.11072v2 Announce Type: replace Abstract: Simultaneous speech translation requires translating source speech into a target language in real-time while handling non-monotonic word dependencie

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

Single-Teacher View Augmentation: Enhancing Knowledge Distillation with Student-Guided Perturbations

DGX agent

arXiv:2607.11557v2 Announce Type: replace Abstract: Knowledge distillation (KD) typically relies on the fixed perspective of a single teacher, limiting the diversity of supervisory signals. While mult

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD

DGX agent

arXiv:2607.20145v1 Announce Type: cross Abstract: Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed trainin

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Small, Free, and Effective: Orchestrating Open-Weight Small Language Models to Outperform Single LLM for Malware Analysis

DGX agent

arXiv:2607.20216v1 Announce Type: cross Abstract: Malware analysis demands rapid interpretation of complex detonation reports spanning filesystem, network, and process behaviours. While large language

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Solar Open 2 Technical Report

DGX agent

arXiv:2607.20062v1 Announce Type: new Abstract: We present Solar Open 2, a 250B-A15B Mixture-of-Experts language model built for long-horizon agentic tasks, scaled up from Solar Open 1 (Solar Open 100

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

Sophisticated Policies from Epistemic Priors

DGX agent

arXiv:2607.19518v1 Announce Type: new Abstract: Sophisticated Inference is a variant of active inference often associated with recursive belief modeling and tree search. We argue that its central comp

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Spectral-LSH: Sub-Quadratic Prompt Compression via Krylov-Projected Locality-Sensitive Hashing

DGX agent

arXiv:2607.19368v1 Announce Type: new Abstract: Long-prompt inference remains expensive because prefill attention scales quadratically with sequence length. We propose Spectral-LSH, a training-free pr

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Start Customizing NVIDIA Nemotron 3 Nano with Prime Intellect Lab in Minutes

DGX agent

Prime Intellect Lab offers a streamlined, hosted reinforcement‑learning workflow that lets users customize the NVIDIA Nemotron 3 Nano in minutes. The process establishes a baseline, trains the model o

model-releasesnvidia-developer
23 Jul 2026
Model Releases

Stateful Guardrails for Multi-Turn LLM Systems: A Conversational Risk Accumulation Framework

DGX agent

arXiv:2607.19361v1 Announce Type: cross Abstract: Most safety guardrails for large language models (LLMs) evaluate each prompt-response pair in isolation, which misses failures that arise only over a

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Statistical Inference for Rank Allocation in Low-Rank Adaptation

DGX agent

arXiv:2607.20205v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has become a widely used parameter-efficient fine-tuning method for large language models. Since different modules and laye

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

Statistically Grounded Sparse-Feature Interventions for Activation-Space Control in Large Language Models

DGX agent

arXiv:2607.19364v1 Announce Type: new Abstract: Activation steering offers a lightweight alternative to fine-tuning for behavioral control of large language models, but SAE-based steering methods ofte

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

STN-TGAT: Top-K Portfolio Construction via Prior-Guided Graph Attention with Learnable Soft-Threshold Sparsification

DGX agent

arXiv:2607.19385v1 Announce Type: new Abstract: This paper tackles the problem of stock ranking and portfolio construction under realistic investment settings by jointly modeling temporal dynamics and

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

Strength-Parity Ensembling with Parameter-Isolated Experts for Multi-Task Affect Recognition

DGX agent

arXiv:2607.16290v2 Announce Type: replace Abstract: Leading entries on the multi-task track of the 11th ABAW challenge rely on heavy ensembling, yet which member is worth adding to an already strong e

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

StrokeSeg2: Stroke Lesion Segmentation in Clinical Research Workflows

DGX agent

arXiv:2607.19901v1 Announce Type: new Abstract: Deep learning frameworks like nnU-Net achieve state-of-theart brain lesion segmentation performance but remain difficult to deploy in clinical research

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

SUM: Unified Geometric Surgery on Spatio-Temporal Adaptation Vectors for Federated Class Incremental Learning

DGX agent

arXiv:2607.19384v1 Announce Type: new Abstract: Real-world intelligent systems often require both distributed collaboration across data-isolated clients and continual adaptation to evolving tasks. Thi

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

SynGallery: A Synthetic Gallery of Real Paintings for Instance-Level Artwork Recognition

DGX agent

arXiv:2607.18907v1 Announce Type: new Abstract: Instance-level artwork recognition requires matching a handheld visitor photograph to a specific work in a large museum collection. This is challenging

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models

DGX agent

arXiv:2607.19608v1 Announce Type: new Abstract: Instruction tuning is meant to make language models follow user requests, yet it is unclear whether small models comply when an instruction conflicts wi

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Convergent Category Geometry in Small Language Models

DGX agent

arXiv:2607.16741v2 Announce Type: replace Abstract: Burger et al. (2024) demonstrated that truth representations in large language models are universal across statement polarity but reside within a mu

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

The Blessing of Dimensionality: How Near-Orthogonality in High-Dimensional Spaces Explains Temporal Portability

DGX agent

arXiv:2607.20301v1 Announce Type: cross Abstract: Fine-tuning has been widely used to adapt large language models (LLMs) for domain-specific tasks. Parameter efficient fine-tuning (PEFT) methods such

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

The Blueprint: How Voicify makes AI-enabled ordering a delight for customers

DGX agent

Welcome to The Blueprint, a new feature where we highlight how Google Cloud customers are tackling unique and common challenges across industries using the latest AI and cloud technologies. We hope to

model-releasesgoogle-cloud-ai
23 Jul 2026
Model Releases

The Chronos Vulnerability: A Taxonomy of Temporal Persistence and Memory-Based Deception in Agentic AI

DGX agent

arXiv:2607.19433v1 Announce Type: new Abstract: The transition from stateless generative models in artificial intelligence to stateful, autonomous agents represents an architectural evolution that, wh

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

The first known runaway AI agent - or a very bad marketing stunt?

DGX agent

The first known runaway AI agent - or a very bad marketing stunt? Martin Alderson's commentary on the OpenAI accidental cyberattack against Hugging Face includes a couple of details I hadn't considere

model-releasessimon-willison
23 Jul 2026
Model Releases

The first router built for generative media is here.

DGX agent

The first router built for generative media is here. Runway launches AI model router as generative media gets crowded https://techcrunch.com/2026/07/23/runway-bets-on-ai-model-routing-as-generative-me

model-releasescristobal-valenzuela--x
23 Jul 2026
Model Releases

The Maskability Index: Predicting Task-Objective Alignment in Pretrained Language Models

DGX agent

arXiv:2607.20265v1 Announce Type: cross Abstract: Large-scale pretrained language models such as T5 and BERT have demonstrated strong capabilities for generating structured knowledge. However, their p

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

The World Model Remembers, the Actor Forgets: Dream Rehearsal for Continual Model-Based RL

DGX agent

arXiv:2607.19749v1 Announce Type: cross Abstract: Model-based reinforcement-learning agents of the DreamerV3 family forget catastrophically when trained on task sequences, even when an unbounded repla

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Think Sparse, Predict Dense: Continuous Thought Machines for Image Super-Resolution

DGX agent

arXiv:2607.18856v1 Announce Type: new Abstract: Continuous Thought Machines introduce an internal temporal dimension in which neuron-level histories and synchronization-derived representations evolve

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation

DGX agent

arXiv:2509.24739v4 Announce Type: replace Abstract: Vision-Language Foundation Models (VLMs), trained on large-scale multimodal datasets, have driven significant advances in Artificial Intelligence (A

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Toward Anthropomorphic Dialogue: A Closed-Loop Framework for Human-Like Chat Generation, Evaluation, and Preference Alignment

DGX agent

arXiv:2607.17191v2 Announce Type: replace Abstract: Human-like private chat requires more than fluent response generation: a system must preserve persona, relationship, memory, bounded knowledge, medi

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Train the Model, Not the Reader: Decodability Supervision for Verifiable Activation Explanations

DGX agent

arXiv:2607.20379v1 Announce Type: new Abstract: Natural-language autoencoders score explanations of hidden activations by reconstruction: an explanation is deemed faithful if the activation can be reg

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Trained a 32B FLUX.2 LoRA on a 24GB AMD 7900 XTX, native ROCm on Windows — full guide + patches

DGX agent

TL;DR: Everyone says QLoRA past ~13B is dead on a 24GB card. I got the full 32B FLUX.2 dev transformer QLoRA-training resident on the GPU on a 7900 XTX under native ROCm on Windows (no ZLUDA, no CUDA

model-releasesr-stablediffusion
23 Jul 2026
Model Releases

Trend strength predicts when generative foundation models win: a power-controlled benchmark, a mechanism, and an actionable selection rule

DGX agent

arXiv:2607.19383v1 Announce Type: cross Abstract: Pretrained generative foundation models cast forecasting as conditional generation from a learned predictive distribution and forecast unseen series z

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

TriAgent: Divergence-Aware Multi-Agent Committees for Cost-Efficient Financial Sentiment Analysis

DGX agent

arXiv:2607.19794v1 Announce Type: new Abstract: Production LLM-based financial sentiment analysis faces a structural cost trap: most queries are trivially classifiable, yet expensive cloud reasoners p

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

Trusted Multi-View Deep Learning Classification of Fetal Congenital Heart Disease with Feature-level and Decision-level Fusion

DGX agent

arXiv:2606.15265v2 Announce Type: replace Abstract: Congenital heart disease (CHD) refers to the abnormal anatomical structure caused by the abnormal development of the heart and great vessels during

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Understanding Generative AI-mediated User Engagement with Academic Library Resources

DGX agent

arXiv:2607.20328v1 Announce Type: cross Abstract: This study empirically analyzed generative AI as an emerging discovery pathway to academic library resources. Utilizing web analytics from August 2023

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Unified Prediction and Planning via Conflict-Aware Disjoint Parameter Training

DGX agent

arXiv:2607.19971v1 Announce Type: new Abstract: Accurate motion prediction of surrounding agents and safe motion planning are two closely coupled key tasks for social robot navigation in crowded envir

model-releasesarxiv-cs-ro
23 Jul 2026
Model Releases

Universality Reconsidered: Rethinking the Validation of Foundation Models for General-Purpose 3D Medical Segmentation

DGX agent

arXiv:2602.07643v2 Announce Type: replace Abstract: Foundation models have emerged as a transformative paradigm in 3D medical imaging, with the promise of unified quantitative analysis across diverse

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Unlearning as Distribution Restoration: A Controlled Counterfactual Study, a Validated Selective Screen, and the Limits of Oracle-Free Certification

DGX agent

arXiv:2607.19442v1 Announce Type: cross Abstract: Machine unlearning is commonly evaluated by matching a retrained oracle on trained probes. In a controlled nonce-fact testbed with a matched retrainin

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

v0.32.3

DGX agent

What's Changed mlx update by @dhiltgen in #17332 model/parsers: finalize incomplete GLM tool calls by @dhiltgen in #17250 docs: update retirements by @mxyng in #17289 model: align Laguna with upstream

model-releasesollama-releases
23 Jul 2026
Model Releases

VG3S: Visual Geometry Grounded Gaussian Splatting for Semantic Occupancy Prediction

DGX agent

arXiv:2603.06210v2 Announce Type: replace-cross Abstract: 3D semantic occupancy prediction has become a crucial perception task for comprehensive scene understanding in autonomous driving. While recen

model-releasesarxiv-cs-ro
23 Jul 2026
Model Releases

We built this experience based on feedback from early testers and physicians. With your permission, ChatGPT can use relevant context you’ve …

DGX agent

We built this experience based on feedback from early testers and physicians. With your permission, ChatGPT can use relevant context you’ve connected in Health across your conversations — to help you

model-releasesopenai--x
23 Jul 2026
Model Releases

When Does Knowledge Distillation Hurt? Reliability-Aware Distillation for Low-Resource Language Summarization

DGX agent

arXiv:2607.19956v1 Announce Type: cross Abstract: Knowledge distillation (KD) is a standard approach for compressing sequence-to-sequence models, but its per-sample effects are rarely examined. On the

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

When Shippers Become Algorithms: Candidate Exposure, Information Design, and the Concentration of LLM-Mediated Freight Markets

DGX agent

arXiv:2607.19967v1 Announce Type: cross Abstract: Shippers are beginning to delegate carrier selection to large language model (LLM) agents. We ask what such delegation does to a freight matching mark

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

When Visual Evidence is Ambiguous: Pareidolia as a Diagnostic Probe for Vision Models

DGX agent

arXiv:2603.03989v3 Announce Type: replace-cross Abstract: When visual evidence is ambiguous, vision models must decide how to interpret face-like patterns. Face pareidolia, the perception of faces in

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

WHU-PCPR: A cross-platform heterogeneous point cloud dataset for place recognition in complex urban scenes

DGX agent

arXiv:2601.06442v2 Announce Type: replace Abstract: Point Cloud-based Place Recognition (PCPR) demonstrates considerable potential in applications such as autonomous driving, robot localization and na

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

WorldPack: Dynamic Frame Compression for Long-context Video World Modeling

DGX agent

arXiv:2512.02473v2 Announce Type: replace-cross Abstract: Video world models have attracted significant attention for their ability to produce high-fidelity future visual observations conditioned on p

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories

DGX agent

arXiv:2607.15330v2 Announce Type: replace Abstract: We present Xiaomi-Robotics-1, a foundational vision-language-action (VLA) model capable of (1) following diverse language instructions to perform a

model-releasesarxiv-cs-ro
23 Jul 2026
Model Releases

You can also use ChatGPT Voice in Codex from the iOS app with paired remote access. Android support is coming soon.

DGX agent

OpenAI has released ChatGPT Voice for the desktop app, enabling users to control their computer and direct multiple agents running in ChatGPT Work or Codex using voice commands. The feature, powered b

model-releasesopenai--x
23 Jul 2026
← Previous
1…9596979899…470
Next →