AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlog
88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,499 results
Model Releases

Bittensor Agent Arenas as a Trajectory Primitive: Distilling a Shopping Agent from ShoppingBench Subnet Traces

DGX agent

arXiv:2606.10064v1 Announce Type: cross Abstract: Small-model agentic post-training is bottlenecked less by the algorithm than by the trajectory substrate it consumes. Leading recipes (RLVR, group-rel

model-releasesarxiv-cs-ai
10 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Claude Fable won’t answer basic biology questions

DGX agent

Anthropic just released Claude Fable 5, calling it the most powerful AI model it has ever made widely available and praising its skills in biology, among others. But the model won't answer basic biolo

model-releasesthe-verge-ai
10 Jun 2026
Model Releases

CLP: Collocation-Length Prediction for Zero-Loss Adaptive Multi-Token Inference

DGX agent

arXiv:2606.10935v1 Announce Type: cross Abstract: Large language model inference is bottlenecked by autoregressive decoding, where each token requires a full forward pass. Multi-token prediction (MTP)

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Continual LLM Upcycling: A Predictor-Gated Bank-Wise Sparsity Training Recipe for Dense-to-Sparse LLMs

DGX agent

arXiv:2606.10722v1 Announce Type: new Abstract: We study dense-to-sparse continual training as a way to construct channel-sparse large language models from dense checkpoints. Starting from a Qwen2.5-8

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Disjoint or Overlapping? Inference Windowing for Reconstruction-Based Time Series Anomaly Detection

DGX agent

arXiv:2606.09874v1 Announce Type: new Abstract: Reconstruction-based methods are widely used for time series anomaly detection, where models are trained to reconstruct subsequences, and anomalies are

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Don't waste SAM

DGX agent

arXiv:2606.10696v1 Announce Type: new Abstract: Meta AI has recently released the Segment Anything Model (SAM), which demonstrates exceptional zero-shot image segmentation performance across various t

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Efficient RWKV-based Representation Learning for 3D Point Clouds

DGX agent

arXiv:2606.10395v1 Announce Type: new Abstract: The recent receptance weighted key value (RWKV) model combines RNN-style recurrence, offering a linear-complexity alternative to Transformers' quadratic

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Hyperparameter Learning for Latent Factorization of Tensors for Representation Learning to Large-scale Dynamic Weighted Directed Network

DGX agent

arXiv:2606.09880v1 Announce Type: new Abstract: Large-scale dynamic weighted directed networks (DWDNs) are widely used to model time-varying interactions among nodes. Latent factorization of tensors (

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Monte Carlo Pass Search: Using Trajectory Generation for 3D Counterfactual Pass Evaluation in Football

DGX agent

arXiv:2606.11120v1 Announce Type: new Abstract: We recast pass evaluation in football (soccer) as a Monte Carlo Tree Search (MCTS)-like evaluation problem whose components mostly exist in the literatu

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

One Step Closer to Ground Truth: A Multi-Scale Residual-Aware Representation Learning Pipeline for Predicting Time Series Data

DGX agent

arXiv:2606.10678v1 Announce Type: new Abstract: Transformer-based models have emerged as leading paradigms in time-series forecasting in recent years, employing self-attention mechanisms to capture lo

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

PreAct-Bench: Benchmarking Predictive Monitoring in LLMs

DGX agent

arXiv:2606.09890v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents capable of executing multi-step action trajectories toward a given objecti

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Quo Vadis, Visual In-Context Learning? A Unified Benchmark Across Domains and Tasks

DGX agent

arXiv:2606.10967v1 Announce Type: new Abstract: Visual in-context learning has been proposed as a pathway towards dynamic models that can generate predictions based on a provided context and thereby c

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

DGX agent

arXiv:2510.14828v3 Announce Type: replace Abstract: Improving the reasoning capabilities of embodied agents is crucial for robots to complete complex human instructions in long-view manipulation tasks

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SAFE: An LLM-as-Verifier Framework for Evidence-Grounded Multi-Hop Reasoning

DGX agent

arXiv:2604.01993v2 Announce Type: replace-cross Abstract: Multi-hop QA benchmarks often reward Large Language Models (LLMs) for spurious correctness, where models reach correct answers through invalid

model-releasesarxiv-cs-ai
10 Jun 2026
Applications

Segmentation-Driven Monocular Shape from Polarization based on Physical Model

DGX agent

arXiv:2601.04776v2 Announce Type: replace Abstract: Monocular shape-from-polarization (SfP) leverages the intrinsic relationship between light polarization properties and surface geometry to recover s

applicationsarxiv-cs-cv
10 Jun 2026
Research

Superficial Beliefs in LLM Decision-Making

DGX agent

arXiv:2606.11016v1 Announce Type: new Abstract: We ask whether large language models (LLMs) merely imitate rationales when choosing between two options, or whether their choices reflect a systematic u

researcharxiv-cs-ai
10 Jun 2026
Model Releases

U-TTT: Towards Generalizable PET Image Denoising via Test-Time Training

DGX agent

arXiv:2606.11032v1 Announce Type: new Abstract: Existing deep learning models for Positron Emission Tomography (PET) image denoising often suffer from severe performance degradation under distribution

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Validation-Stage Combinatorial Fusion Analysis for Imbalanced Credit-Card Fraud Detection

DGX agent

arXiv:2606.10393v1 Announce Type: new Abstract: Credit-card fraud detection is difficult because fraudulent transactions are rare, costly, and unevenly distributed. Strong gradient-boosted tree models

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

WebChallenger: A Reliable and Efficient Generalist Web Agent

DGX agent

arXiv:2606.10423v1 Announce Type: new Abstract: Autonomous web navigation remains challenging for LLM agents, and the strongest generalist systems rely on proprietary reasoning models whose inference

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

What Fits (Into Few Tokens) Doesn't Overfit: Compression and Generalization in ML Research Agents

DGX agent

arXiv:2606.11045v1 Announce Type: new Abstract: Reusing a held-out benchmark adaptively should, in principle, invite overfitting. Yet benchmark-driven machine learning (ML) has produced surprisingly l

model-releasesarxiv-cs-ai
10 Jun 2026
Research

A Machine Learning-Enhanced Hopf-Cole Formulation for Nonlinear Gas Flow in Porous Media

DGX agent

arXiv:2603.11250v2 Announce Type: replace-cross Abstract: Accurate modeling of gas flow through porous media is critical for many technological applications, including reservoir performance prediction

researcharxiv-cs-lg
9 Jun 2026
Safety

An Agency-Transferring Model-Free Policy Enhancement Technique

DGX agent

arXiv:2606.09825v1 Announce Type: cross Abstract: Training reinforcement learning (RL) policies from scratch is costly: it requires careful reward and environment design, extensive tuning, and substan

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs

DGX agent

arXiv:2606.07643v1 Announce Type: cross Abstract: Recent advances in Omni-Multimodal Large Language Models (Omni-MLLMs) have enabled strong integration of vision, audio, and language. However, their a

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Beyond Pass Rate: A Multilingual, Execution-Grounded Evaluation of Open Code LLMs

DGX agent

arXiv:2606.08840v1 Announce Type: new Abstract: Code generation models are typically compared using compact execution benchmarks and aggregate pass rates, but such summaries obscure how performance va

model-releasesarxiv-cs-ai
9 Jun 2026
Research

CRAG: Can 3D Generative Models Help 3D Assembly?

DGX agent

arXiv:2602.22629v2 Announce Type: replace Abstract: Most existing 3D assembly methods treat the problem as pure pose estimation, rearranging observed parts via rigid transformations. In contrast, huma

researcharxiv-cs-cv
9 Jun 2026
Model Releases

CRANE: Knowledge Editing for Reasoning MLLMs

DGX agent

arXiv:2606.09033v1 Announce Type: new Abstract: The emergence of reasoning multimodal large language models (MLLMs), which generate explicit chain-of-thought (CoT) reasoning before producing answers,

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Decomposable Neuro Symbolic Regression

DGX agent

arXiv:2511.04124v3 Announce Type: replace Abstract: Symbolic regression (SR) models complex systems by discovering mathematical expressions that capture underlying relationships in observed data. Howe

model-releasesarxiv-cs-lg
9 Jun 2026
Industry

Defend against frontier cyber models: Cloudflare's architecture as customer zero

DGX agent

In our post about Project Glasswing, we made the argument that the architecture around a vulnerability matters more than the speed of the patch. Here we walk through what that architecture looks like,

industrycloudflare-ai
9 Jun 2026
Model Releases

Defending Against Malicious Finetuning by Scaling Train-time Adversarial Attacks

DGX agent

arXiv:2606.07970v1 Announce Type: cross Abstract: Current open-weight large language models (LLMs) are prone to malicious finetuning attacks, which could compromise the safety alignment of LLMs with o

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

EditSSC: Toward Editable Semantic Occupancy Scenes with Unconditional Diffusion Models

DGX agent

arXiv:2606.09273v1 Announce Type: new Abstract: 3D semantic scene generation is crucial for autonomous driving applications, yet most methods rely on complex 3D-specific architectures such as triplane

agentsarxiv-cs-cv
9 Jun 2026
Model Releases

Google announces Gemini 3.5 Live Translate for instant voice-to-voice translation

DGX agent

Gemini 3.5 Live Translate is Google's latest audio model delivering near real-time speech-to-speech translation in over 70 languages. The model automatically detects 70+ languages and generates smooth

model-releasesars-technica
9 Jun 2026
Model Releases

How Much Capacity Does EEG Denoising Need? Ultra-Compact Networks reveal Benchmark Saturation and Metric-Utility Gap

DGX agent

arXiv:2606.08594v1 Announce Type: new Abstract: Deep learning EEG denoising architectures have scaled from tens of thousands to tens of millions of parameters, yet no prior study has isolated model ca

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation

DGX agent

arXiv:2601.04498v2 Announce Type: replace-cross Abstract: Infographics are composite visual artifacts that combine data visualizations with textual and illustrative elements to communicate information

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

PEDRA: Evaluating the Realism of Pedestrian Dynamics in Video Generation

DGX agent

arXiv:2510.20182v2 Announce Type: replace Abstract: Pedestrian simulation traditionally relies on expert-tuned, hand-crafted models that limit scalability and generalization. Meanwhile, large-scale vi

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Petri Net Modeling and Deadlock-Free Scheduling of Attachable Heterogeneous AGV Systems

DGX agent

arXiv:2508.00724v2 Announce Type: replace-cross Abstract: The increasing demand for flexible automation has accelerated the adoption of heterogeneous automated guided vehicles (AGVs). This work invest

safetyarxiv-cs-ro
9 Jun 2026
Model Releases

PLAGUE: Plug-and-play framework for Lifelong Adaptive Generation of Multi-turn Exploits

DGX agent

arXiv:2510.17947v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are improving at an exceptional rate. With the advent of agentic workflows, multi-turn dialogue has become the de

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks

DGX agent

arXiv:2606.07968v1 Announce Type: cross Abstract: Reasoning-capable large language models can be induced to spend their generation budget on injected decoy tasks rather than answering the user's quest

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Scaling Laws for Masked-Reconstruction Transformers on Single-Cell Transcriptomics

DGX agent

arXiv:2602.15253v2 Announce Type: replace Abstract: Neural scaling laws -- power-law relationships between loss, model size, and data -- have been extensively documented for language and vision transf

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Sci-Rho: A Multilingual Visually-Grounded Symbolic Benchmark for STEM Problems

DGX agent

arXiv:2606.08034v1 Announce Type: cross Abstract: Symbolic benchmarks have emerged as a key approach to assess model robustness under minor modifications to STEM-related questions. However, existing s

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Subtitle-Aligned Fine-Tuning of Whisper for Swiss German ASR: Benchmark Contamination, Convention Mismatch, and an Honest Baseline at 25.6% WER (13.8% cWER)

DGX agent

arXiv:2606.07608v1 Announce Type: cross Abstract: We present a systematic study of fine-tuning OpenAI's Whisper large-v3 for Swiss German ASR, using 1,367 hours of broadcast speech paired with Standar

model-releasesarxiv-cs-ai
9 Jun 2026
Research

TLDR: Compressing Audio Tokens for Efficient Autoregressive Text-to-Speech

DGX agent

arXiv:2606.09019v1 Announce Type: cross Abstract: Codec-based autoregressive (AR) speech language models have achieved strong text-to-speech (TTS) quality by modeling speech as sequences of discrete a

researcharxiv-cs-ai
9 Jun 2026
Model Releases

VATS: Exploiting Implicit Authority in Error-Path Injection via Systematic Mutation

DGX agent

arXiv:2606.07992v1 Announce Type: new Abstract: As the Model Context Protocol (MCP) standardizes tool-calling for autonomous agents, it introduces a critical, unexamined attack surface: the error-hand

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning?

DGX agent

arXiv:2606.07872v1 Announce Type: new Abstract: When a multimodal large language model answers a visual reasoning question correctly, is the prediction actually supported by the task-critical visual e

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Your Self-Play Algorithm is Secretly an Adversarial Imitator: Understanding LLM Self-Play through the Lens of Imitation Learning

DGX agent

arXiv:2602.01357v2 Announce Type: replace Abstract: Self-play post-training methods has emerged as an effective approach for finetuning large language models and turn the weak language model into stro

safetyarxiv-cs-lg
9 Jun 2026
Model Releases

Zero-Shot Learning in Industrial Scenarios: New Large-Scale Benchmark, Challenges and Baseline

DGX agent

arXiv:2606.07965v1 Announce Type: new Abstract: Large Visual Language Models (LVLMs) have achieved remarkable success in vision tasks. However, the significant differences between industrial and natur

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ZIPP:Zero-shot Image Personalization from Personas

DGX agent

arXiv:2606.08841v1 Announce Type: new Abstract: Text-to-image diffusion models are increasingly deployed in open-ended creative contexts, yet their outputs remain impersonal, optimized for aggregate a

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

CrowdMath: A Dataset of Crowdsourced Mathematical Research Discussions

DGX agent

arXiv:2606.06526v1 Announce Type: new Abstract: Large language models have made substantial progress on mathematical reasoning, but existing benchmarks typically evaluate well-specified problems with

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Entropy as a Structural Prior: How a Log-Barrier on DiT Belief Space Drives Musical Diversity and Development

DGX agent

arXiv:2606.07207v1 Announce Type: cross Abstract: Confidence-based loss weighting is usually avoided in generative models because it accelerates errors when the model is confidently wrong, but this in

model-releasesarxiv-cs-lg
8 Jun 2026
← Previous
1…357358359360361…1323
Next →