AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Local Ai

Trace, Verify, and Correct: A Training-Free Framework for Spatial Reasoning in Multimodal LLMs

DGX agent

arXiv:2608.04759v1 Announce Type: cross Abstract: Although Multimodal Large Language Models (MLLMs) have made substantial progress, their spatial reasoning may still produce intermediate judgments inc

local-aiarxiv-cs-cl
6 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Tropical Algebraic Geometry for Neuronal Representations: An Arakelov-Green Measure Based Descriptor for Graph Learning

DGX agent

arXiv:2608.04460v1 Announce Type: cross Abstract: The quantitative analysis of 3D neuronal morphologies requires capturing both graph topology and spatial geometry. Current message-passing Graph Neura

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Two-Level Meta-Rubrics for Evaluating Open-Ended Generation: GAMUT, a Benchmark for Factual Completeness

DGX agent

arXiv:2607.19322v2 Announce Type: replace Abstract: Rubric-based evaluation of open-ended generation faces a fundamental tension between expressiveness and reliability. Authoring a faithful rubric req

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

UG-UMRE: Uncertainty-Guided Modality Augmentation and Distributional Calibration for Unified Multimodal Relation Extraction

DGX agent

arXiv:2608.04949v1 Announce Type: cross Abstract: Unified Multimodal Relation Extraction (UMRE) aims to identify intra-modal and cross-modal relations between textual entities and visual objects. Howe

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Visual Representation Matters: Exploiting Temporal Differences in Video-to-Audio Generation

DGX agent

arXiv:2608.04902v1 Announce Type: new Abstract: Video-to-audio (V2A) generation extends image-to-audio generation (I2A) by introducing consecutive frames that provide essential temporal cues for audio

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Your agentic summer: No-cost lessons from Google experts to build and scale agents

DGX agent

I’ve talked to developers, IT leaders, and builders who all ask the same question: How do we actually get agents into production? The answer isn't theoretical — it's hands-on. Whether it’s designing a

model-releasesgoogle-cloud-ai
6 Aug 2026
Model Releases

Zero-shot reasoning for simulating scholarly peer-review

DGX agent

arXiv:2510.02027v2 Announce Type: replace Abstract: Scholarly publishing requires scalable scrutiny supported by auditable evidence. This paper presents a two-component benchmark of xPeer, the peer-re

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

3DGSI-Assessor: A Large-Scale Dataset and An LMM-based Method for 3D Gaussian Splatting Image Quality Assessment

DGX agent

arXiv:2608.03279v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has become a dominant representation for real-time novel view synthesis (NVS), yet its storage footprint makes compression

model-releasesarxiv-cs-cv
5 Aug 2026
Hardware

A Deployment-Friendly Foundational Framework for Efficient Computational Pathology

DGX agent

arXiv:2602.14010v2 Announce Type: replace-cross Abstract: Pathology foundation models (PFMs) generalize well across computational pathology tasks but remain costly for gigapixel whole-slide image anal

hardwarearxiv-cs-ai
5 Aug 2026
Hardware

AcceptMoE: Commitment-Weighted Self-Sizing Verifier Expert Sets for Efficient MoE Speculative Decoding

DGX agent

arXiv:2608.02989v1 Announce Type: cross Abstract: Speculative decoding verifies a tree of draft tokens in one target-model forward pass. For a mixture-of-experts (MoE) target, however, parallel verifi

hardwarearxiv-cs-cl
5 Aug 2026
Research

Adversarial Purification by Consistency-aware Latent Space Optimization on Data Manifolds

DGX agent

arXiv:2412.08394v2 Announce Type: replace Abstract: Deep neural networks (DNNs) are vulnerable to adversarial samples crafted by adding imperceptible perturbations to clean data, potentially leading t

researcharxiv-cs-lg
5 Aug 2026
Model Releases

Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation

DGX agent

arXiv:2608.03166v1 Announce Type: new Abstract: Role-Playing Language Agents (RPLAs) are increasingly deployed in high-stakes applications such as healthcare assistance, customer support, and educatio

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Aligned in Form, Not in Meaning: The Comprehension - Containment Decoupling of LLM Safety in Low-Resource Bangla Derogatory Speech

DGX agent

arXiv:2608.02941v1 Announce Type: new Abstract: We audit five frontier large language models on native Bangla derogatory speech (gali) across six protocols to test a single hypothesis: Comprehension-C

safetyarxiv-cs-cl
5 Aug 2026
Applications

Amortized Interventional Forecasting for Multivariate CIR Processes

DGX agent

arXiv:2608.03715v1 Announce Type: new Abstract: Mean-reverting dynamics are pervasive in finance, and the Cox--Ingersoll--Ross (CIR) process is a standard model for the time series they produce, from

applicationsarxiv-cs-lg
5 Aug 2026
Model Releases

Anyone interested in building a harness-only benchmark?

DGX agent

There are a lot of LLM benchmarks but few, if any, harness benchmarks. I am thinking this would be a really good community project to build one. End goal: a leaderboard of harness performance (multipl

model-releasesr-localllama
5 Aug 2026
Model Releases

ATFlash: Per-RoPE-Wavelength Attention Windows for Compute/Memory-Efficient LLM Inference

DGX agent

arXiv:2608.02947v1 Announce Type: cross Abstract: The attention score with rotary position embeddings (RoPE) decomposes exactly into a sum over its 2D-rotation frequency pairs, and each pair's wavelen

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

BBOWP-Bench: Evaluating LLMs on Black-Box Optimization Word Problems

DGX agent

arXiv:2608.02612v1 Announce Type: new Abstract: Formulating an optimization problem strongly affects the quality of the final solution, yet good formulations usually require substantial expertise. Rec

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Beyond Either-Or Reasoning: Transduction and Induction as Cooperative Problem-Solving Paradigms

DGX agent

arXiv:2505.14744v3 Announce Type: replace-cross Abstract: Traditionally, in Programming-by-example (PBE) the goal is to synthesize a program from a small set of input-output examples. Lately, PBE has

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

CADET: Physics-Grounded Causal Auditing and Training-Free Deconfounding of End-to-End Driving Planners

DGX agent

arXiv:2606.14438v3 Announce Type: replace-cross Abstract: End-to-end (E2E) autonomous-driving planners trained by imitation are prone to statistical shortcuts: they associate scene elements that merel

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Channel-wise Dynamic Knowledge Distillation via Adaptive Sample Generation for Action Recognition

DGX agent

arXiv:2608.03100v1 Announce Type: new Abstract: Knowledge Distillation (KD) offers a promising yet underexplored path for compressing large action recognition models. However, existing KD methods suff

safetyarxiv-cs-cv
5 Aug 2026
Model Releases

Cloudflare launches Identity-Aware AI Gateway to track who is using AI

DGX agent

Cloudflare Inc. today launched Identity-Aware AI Gateway, a service that attaches a verified identity to every artificial intelligence request leaving a company network. Information technology and sec

model-releasessiliconangle
5 Aug 2026
Agents

ContinualSkillBench: Can LLM Agents Truly Evolve Their Capabilities?

DGX agent

arXiv:2608.03874v1 Announce Type: new Abstract: Modern agent frameworks equip large language models with external skill libraries to solve complex tasks. However, it remains unclear whether these syst

agentsarxiv-cs-ai
5 Aug 2026
Applications

Cross-Country Learning for National Infectious Disease Forecasting Using European Data

DGX agent

arXiv:2601.20771v2 Announce Type: replace-cross Abstract: Accurate forecasting of infectious disease incidence is critical for public health planning and timely intervention. While most data-driven fo

applicationsarxiv-cs-lg
5 Aug 2026
Applications

CURV: Enhancing Chart Understanding Through Curriculum Visual Grounded Reasoning

DGX agent

arXiv:2608.02833v1 Announce Type: cross Abstract: Chart question answering (CQA) requires multimodal large language models (MLLMs) to integrate visual comprehension with logical reasoning, yet current

applicationsarxiv-cs-ai
5 Aug 2026
Safety

CVPO: Enhancing LLM Reinforcement Learning Reasoning via Value-Variance Adaptation and Dynamic Curriculum Learning

DGX agent

arXiv:2608.03068v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as an effective method for enhancing the reasoning capabilities of large language models (LLMs). However, exis

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

DAIF: A Data-Driven Intermediate Fusion Framework for Multimodal Supervised Learning via Approximate Message Passing

DGX agent

arXiv:2608.02769v1 Announce Type: cross Abstract: Multimodal supervised learning seeks to leverage multiple heterogeneous data sources to improve predictive performance. A central challenge is determi

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces

DGX agent

arXiv:2608.03451v1 Announce Type: new Abstract: Data agents enable natural-language analytics over organizational workspaces, where relevant evidence may be scattered across databases, structured file

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Deepseek V4 Flash just hit Colibri, does anyone have numbers?

DGX agent

I'm mosty interested in 128-192GB VRAM with 128-256GB RAM to spare, so SSD streaming is basically not even necessary. Seems only FP4 is supported, so older hardware will likely be slow - no Unsloth GG

model-releasesr-localllama
5 Aug 2026
Model Releases

Diversity is Not Ambiguity: Toward Accurate and Efficient Ambiguity Detection for Open-Domain QA

DGX agent

arXiv:2608.03177v1 Announce Type: new Abstract: How can question answering (QA) systems determine whether a query is ambiguous? Ambiguity detection is essential in open-domain QA, as misclassification

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

EditFlow3D: Automated Local Editing of 3D Assets with Trajectory Preservation

DGX agent

arXiv:2608.03179v1 Announce Type: new Abstract: Controllable local editing of 3D assets requires precise target localization and appropriate visual guidance. However, existing methods lack a simple ye

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Evaluation Blindness: How Silent Measurement Failures Corrupt AI Systems from Training to Deployment

DGX agent

arXiv:2608.02786v1 Announce Type: new Abstract: AI systems can fail silently. The failure propagates through training loops, evaluation pipelines, and production monitoring stacks until downstream har

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Fail-Fast, Restart-Smart: Early Failure Prediction and Restart for SWE Agentic Tasks

DGX agent

arXiv:2608.03222v1 Announce Type: cross Abstract: Software engineering (SWE) agents resolve repository-level issues through long trajectories that grow increasingly expensive as context accumulates. F

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

FedCritic-MIMO: Communication-Efficient Serverless Federated Critic Learning for Massive-MIMO Resource Control in Open and Disaggregated 6G RANs

DGX agent

arXiv:2608.03852v1 Announce Type: new Abstract: This paper proposes FedCritic-MIMO, a communication-efficient serverless federated multi-agent reinforcement learning framework for AI-native resource c

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Feels like Google could have been the dominating force in AI by open-sourcing the frontier with Gemini, Veo, and Nano Banana. Instead, they …

DGX agent

Feels like Google could have been the dominating force in AI by open-sourcing the frontier with Gemini, Veo, and Nano Banana. Instead, they kept them behind APIs for a few billion dollars in revenue.

model-releasesclem-delangue--x
5 Aug 2026
Model Releases

Forecasting Revenue with its Customer-Base Drivers: When and Why Coordination Helps

DGX agent

arXiv:2608.02911v1 Announce Type: new Abstract: Revenue forecasts guide acquisition budgets, demand planning, and customer-based valuations, yet an aggregate forecast does not show whether change refl

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

... four years later, I got the new Claude Fable 5 to actually build the game https://x.com/simonw/status/2085089518223602058

DGX agent

... four years later, I got the new Claude Fable 5 to actually build the game https://x.com/simonw/status/2085089518223602058 Four years ago today I tweeted about having GPT-3 and DALL-E come up with

model-releasessimon-willison--x
5 Aug 2026
Model Releases

Frozen High-Resolution Inference for Cross-City Object Detection: An AI City Challenge 2026 Study

DGX agent

arXiv:2608.03136v1 Announce Type: new Abstract: Cross-city object detection requires a detector trained in one city to generalize to an unlabeled target city. In AI City Challenge 2026 Track 6, we ana

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

GDPevo: Evaluating Agent Self-Evolution on Real Business Tasks

DGX agent

arXiv:2608.03764v1 Announce Type: new Abstract: Agent self-evolution updates an agent's persistent state from prior experience and reuses it to solve related tasks more effectively. Evaluating self-ev

model-releasesarxiv-cs-ai
5 Aug 2026
Tutorials

GraspMeanFlow: SE(3)-Equivariant MeanFlow for Few-Step 6-DoF Grasp Generation

DGX agent

arXiv:2608.03295v1 Announce Type: new Abstract: Recent data-driven methods for synthesizing 6-DoF grasp poses use generative models to learn complex grasp pose distributions and generate diverse candi

tutorialsarxiv-cs-ro
5 Aug 2026
Model Releases

GUI-Lens: Coarse-to-Fine Cropping for GUI Grounding with General-Purpose VLMs

DGX agent

arXiv:2608.03270v1 Announce Type: cross Abstract: GUI grounding maps natural-language instructions to click locations and is essential for reliable GUI agents. The task remains difficult on high-resol

model-releasesarxiv-cs-ai
5 Aug 2026
Research

HalluTruthQA-4K: A Fine-Grained Corpus and Annotation Process for Arabic Hallucination Detection and Truth Verification

DGX agent

arXiv:2608.03966v1 Announce Type: new Abstract: Large language models can generate fluent Arabic answers while introducing factual errors that are difficult to identify and verify. Existing Arabic hal

researcharxiv-cs-cl
5 Aug 2026
Agents

IR2Solve: Structured Intermediate Representations for Cost-Efficient Optimization Autoformulation

DGX agent

arXiv:2608.02641v1 Announce Type: cross Abstract: Large language models (LLMs) can translate natural-language optimization problems into solver-ready formulations, but direct code generation is brittl

agentsarxiv-cs-ai
5 Aug 2026
Tools

.@Kimi_Moonshot benchmarked K3 endpoints across major inference providers. Together AI leads or ties for #1 on 3 of 4 benchmarks: OCRBench, …

DGX agent

.@Kimi_Moonshot benchmarked K3 endpoints across major inference providers. Together AI leads or ties for #1 on 3 of 4 benchmarks: OCRBench, MMMU Pro Vision, and DeepSWE. Open models like Kimi K3 have

toolstogether-ai--x
5 Aug 2026
Agents

LiveEvalBench: Toward Open-World Evaluation for Web Generation

DGX agent

arXiv:2608.03689v1 Announce Type: new Abstract: Large language models are increasingly capable of synthesizing executable frontend projects, yet existing benchmarks still treat web generation as a sta

agentsarxiv-cs-ai
5 Aug 2026
Hardware

LLM Serving in the Wild: An Empirical Study of Frameworks, Methods, and System Designs

DGX agent

arXiv:2608.03036v1 Announce Type: cross Abstract: Large Language Models (LLMs) are integrated into software systems and AI services, making efficient LLM serving a concern for software engineering. Se

hardwarearxiv-cs-ai
5 Aug 2026
Model Releases

LoCA: Forward-Only LLM Tuning after One-Shot Calibration with Local Credit Assignment

DGX agent

arXiv:2608.03020v1 Announce Type: new Abstract: Parameter-efficient post-training reduces the number of trainable parameters, but still requires repeated end-to-end backpropagation through the frozen

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Maglev: Sliding Recurrent Memory

DGX agent

arXiv:2608.02870v1 Announce Type: new Abstract: We introduce ours{}, a recurrent Transformer architecture with fixed-size memory that generalizes sliding-window attention while remaining parallelizabl

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Measurement Without Validity: The Compounding Reliability Problem in Agentic AI Evaluation

DGX agent

arXiv:2608.00794v2 Announce Type: replace Abstract: Agentic AI evaluation pipelines produce benchmark scores that justify deployment decisions, safety certifications, and regulatory compliance claims.

model-releasesarxiv-cs-ai
5 Aug 2026
← Previous
1…661662663664665…1371
Next →