AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlog
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,676 results
Applications

Cross-Country Learning for National Infectious Disease Forecasting Using European Data

DGX agent

arXiv:2601.20771v2 Announce Type: replace-cross Abstract: Accurate forecasting of infectious disease incidence is critical for public health planning and timely intervention. While most data-driven fo

applicationsarxiv-cs-lg
5 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Applications

CURV: Enhancing Chart Understanding Through Curriculum Visual Grounded Reasoning

DGX agent

arXiv:2608.02833v1 Announce Type: cross Abstract: Chart question answering (CQA) requires multimodal large language models (MLLMs) to integrate visual comprehension with logical reasoning, yet current

applicationsarxiv-cs-ai
5 Aug 2026
Safety

CVPO: Enhancing LLM Reinforcement Learning Reasoning via Value-Variance Adaptation and Dynamic Curriculum Learning

DGX agent

arXiv:2608.03068v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as an effective method for enhancing the reasoning capabilities of large language models (LLMs). However, exis

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

DAIF: A Data-Driven Intermediate Fusion Framework for Multimodal Supervised Learning via Approximate Message Passing

DGX agent

arXiv:2608.02769v1 Announce Type: cross Abstract: Multimodal supervised learning seeks to leverage multiple heterogeneous data sources to improve predictive performance. A central challenge is determi

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces

DGX agent

arXiv:2608.03451v1 Announce Type: new Abstract: Data agents enable natural-language analytics over organizational workspaces, where relevant evidence may be scattered across databases, structured file

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Deepseek V4 Flash just hit Colibri, does anyone have numbers?

DGX agent

I'm mosty interested in 128-192GB VRAM with 128-256GB RAM to spare, so SSD streaming is basically not even necessary. Seems only FP4 is supported, so older hardware will likely be slow - no Unsloth GG

model-releasesr-localllama
5 Aug 2026
Model Releases

Diversity is Not Ambiguity: Toward Accurate and Efficient Ambiguity Detection for Open-Domain QA

DGX agent

arXiv:2608.03177v1 Announce Type: new Abstract: How can question answering (QA) systems determine whether a query is ambiguous? Ambiguity detection is essential in open-domain QA, as misclassification

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

EditFlow3D: Automated Local Editing of 3D Assets with Trajectory Preservation

DGX agent

arXiv:2608.03179v1 Announce Type: new Abstract: Controllable local editing of 3D assets requires precise target localization and appropriate visual guidance. However, existing methods lack a simple ye

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Evaluation Blindness: How Silent Measurement Failures Corrupt AI Systems from Training to Deployment

DGX agent

arXiv:2608.02786v1 Announce Type: new Abstract: AI systems can fail silently. The failure propagates through training loops, evaluation pipelines, and production monitoring stacks until downstream har

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Fail-Fast, Restart-Smart: Early Failure Prediction and Restart for SWE Agentic Tasks

DGX agent

arXiv:2608.03222v1 Announce Type: cross Abstract: Software engineering (SWE) agents resolve repository-level issues through long trajectories that grow increasingly expensive as context accumulates. F

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

FedCritic-MIMO: Communication-Efficient Serverless Federated Critic Learning for Massive-MIMO Resource Control in Open and Disaggregated 6G RANs

DGX agent

arXiv:2608.03852v1 Announce Type: new Abstract: This paper proposes FedCritic-MIMO, a communication-efficient serverless federated multi-agent reinforcement learning framework for AI-native resource c

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Feels like Google could have been the dominating force in AI by open-sourcing the frontier with Gemini, Veo, and Nano Banana. Instead, they …

DGX agent

Feels like Google could have been the dominating force in AI by open-sourcing the frontier with Gemini, Veo, and Nano Banana. Instead, they kept them behind APIs for a few billion dollars in revenue.

model-releasesclem-delangue--x
5 Aug 2026
Model Releases

Forecasting Revenue with its Customer-Base Drivers: When and Why Coordination Helps

DGX agent

arXiv:2608.02911v1 Announce Type: new Abstract: Revenue forecasts guide acquisition budgets, demand planning, and customer-based valuations, yet an aggregate forecast does not show whether change refl

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

... four years later, I got the new Claude Fable 5 to actually build the game https://x.com/simonw/status/2085089518223602058

DGX agent

... four years later, I got the new Claude Fable 5 to actually build the game https://x.com/simonw/status/2085089518223602058 Four years ago today I tweeted about having GPT-3 and DALL-E come up with

model-releasessimon-willison--x
5 Aug 2026
Model Releases

Frozen High-Resolution Inference for Cross-City Object Detection: An AI City Challenge 2026 Study

DGX agent

arXiv:2608.03136v1 Announce Type: new Abstract: Cross-city object detection requires a detector trained in one city to generalize to an unlabeled target city. In AI City Challenge 2026 Track 6, we ana

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

GDPevo: Evaluating Agent Self-Evolution on Real Business Tasks

DGX agent

arXiv:2608.03764v1 Announce Type: new Abstract: Agent self-evolution updates an agent's persistent state from prior experience and reuses it to solve related tasks more effectively. Evaluating self-ev

model-releasesarxiv-cs-ai
5 Aug 2026
Tutorials

GraspMeanFlow: SE(3)-Equivariant MeanFlow for Few-Step 6-DoF Grasp Generation

DGX agent

arXiv:2608.03295v1 Announce Type: new Abstract: Recent data-driven methods for synthesizing 6-DoF grasp poses use generative models to learn complex grasp pose distributions and generate diverse candi

tutorialsarxiv-cs-ro
5 Aug 2026
Model Releases

GUI-Lens: Coarse-to-Fine Cropping for GUI Grounding with General-Purpose VLMs

DGX agent

arXiv:2608.03270v1 Announce Type: cross Abstract: GUI grounding maps natural-language instructions to click locations and is essential for reliable GUI agents. The task remains difficult on high-resol

model-releasesarxiv-cs-ai
5 Aug 2026
Research

HalluTruthQA-4K: A Fine-Grained Corpus and Annotation Process for Arabic Hallucination Detection and Truth Verification

DGX agent

arXiv:2608.03966v1 Announce Type: new Abstract: Large language models can generate fluent Arabic answers while introducing factual errors that are difficult to identify and verify. Existing Arabic hal

researcharxiv-cs-cl
5 Aug 2026
Agents

IR2Solve: Structured Intermediate Representations for Cost-Efficient Optimization Autoformulation

DGX agent

arXiv:2608.02641v1 Announce Type: cross Abstract: Large language models (LLMs) can translate natural-language optimization problems into solver-ready formulations, but direct code generation is brittl

agentsarxiv-cs-ai
5 Aug 2026
Tools

.@Kimi_Moonshot benchmarked K3 endpoints across major inference providers. Together AI leads or ties for #1 on 3 of 4 benchmarks: OCRBench, …

DGX agent

.@Kimi_Moonshot benchmarked K3 endpoints across major inference providers. Together AI leads or ties for #1 on 3 of 4 benchmarks: OCRBench, MMMU Pro Vision, and DeepSWE. Open models like Kimi K3 have

toolstogether-ai--x
5 Aug 2026
Agents

LiveEvalBench: Toward Open-World Evaluation for Web Generation

DGX agent

arXiv:2608.03689v1 Announce Type: new Abstract: Large language models are increasingly capable of synthesizing executable frontend projects, yet existing benchmarks still treat web generation as a sta

agentsarxiv-cs-ai
5 Aug 2026
Hardware

LLM Serving in the Wild: An Empirical Study of Frameworks, Methods, and System Designs

DGX agent

arXiv:2608.03036v1 Announce Type: cross Abstract: Large Language Models (LLMs) are integrated into software systems and AI services, making efficient LLM serving a concern for software engineering. Se

hardwarearxiv-cs-ai
5 Aug 2026
Model Releases

LoCA: Forward-Only LLM Tuning after One-Shot Calibration with Local Credit Assignment

DGX agent

arXiv:2608.03020v1 Announce Type: new Abstract: Parameter-efficient post-training reduces the number of trainable parameters, but still requires repeated end-to-end backpropagation through the frozen

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Maglev: Sliding Recurrent Memory

DGX agent

arXiv:2608.02870v1 Announce Type: new Abstract: We introduce ours{}, a recurrent Transformer architecture with fixed-size memory that generalizes sliding-window attention while remaining parallelizabl

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Measurement Without Validity: The Compounding Reliability Problem in Agentic AI Evaluation

DGX agent

arXiv:2608.00794v2 Announce Type: replace Abstract: Agentic AI evaluation pipelines produce benchmark scores that justify deployment decisions, safety certifications, and regulatory compliance claims.

model-releasesarxiv-cs-ai
5 Aug 2026
Tutorials

MIMIC-MJX: Neuromechanical Emulation of Animal Behavior

DGX agent

arXiv:2511.20532v3 Announce Type: replace-cross Abstract: The primary output of the nervous system is movement and behavior. While recent advances have democratized pose tracking during complex behavi

tutorialsarxiv-cs-ai
5 Aug 2026
Model Releases

Minimax-Optimal Semiparametric Contextual Dynamic Pricing with Multimodal Revenue

DGX agent

arXiv:2608.03142v1 Announce Type: cross Abstract: We study contextual dynamic pricing with arbitrary covariate sequences and bounded, possibly nonbinary purchase quantities. Demand follows a semiparam

model-releasesarxiv-cs-ai
5 Aug 2026
Local Ai

Ollama on mac mini/studio?

DGX agent

I don't own a mac Right now, i want to buy one but its main job will be to host a Ollama or similar software to give me usable AI models in my network, since I don't want to pay for Cloude, GitHub cop

local-air-ollama
5 Aug 2026
Model Releases

On the Implicit Flatness Bias of Sharpness-Aware Minimization: A Linear Stability Analysis with Quantitative Hyperparameter Bounds

DGX agent

arXiv:2608.03197v1 Announce Type: new Abstract: Sharpness-Aware Minimization (SAM) improves generalization by seeking parameters whose loss is robust to local adversarial perturbations, but the quanti

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

On the missing benchmarks layer and a potential solution

DGX agent

arXiv:2608.02996v1 Announce Type: new Abstract: Latin America is missing a foundational layer for native AI development: the benchmark layer. The benchmark layer does two things no other layer can - i

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

One-Point Contraction: Erasing Representational Separability toward Irreversible Deep Forgetting

DGX agent

arXiv:2507.07754v3 Announce Type: replace-cross Abstract: Machine unlearning is usually evaluated by what the classifier outputs: forget-set accuracy, confidence, membership-inference scores. We show

model-releasesarxiv-cs-ai
5 Aug 2026
Industry

OpenAI says the Hugging Face breach involved AI agents creating an internal message board, unnoticed by humans, where they shared exploits and planned the hack (Lily Hay Newman/Wired)

DGX agent

Lily Hay Newman / Wired: OpenAI says the Hugging Face breach involved AI agents creating an internal message board, unnoticed by humans, where they shared exploits and planned the hack — At the Black

industrytechmeme
5 Aug 2026
Research

Output-Aware Rotation for INT2 KV-Cache Quantization

DGX agent

arXiv:2608.02691v1 Announce Type: cross Abstract: The key-value (KV) cache has become a major memory and bandwidth bottleneck in long-context large language model inference, making ultra-low-bit quant

researcharxiv-cs-ai
5 Aug 2026
Safety

Predicting Multilingual Classification and Translation Performance of LLMs with Cross-Lingual Alignment nicode{x2013} Is English Enough?

DGX agent

arXiv:2608.03446v1 Announce Type: new Abstract: Multilingual large language models (LLMs) have been shown to perform better on non-English classification tasks when the representations of the given la

safetyarxiv-cs-cl
5 Aug 2026
Model Releases

Predictive Enhancement Calibration for Latent Breast MRI Virtual Contrast Enhancement

DGX agent

arXiv:2608.03612v1 Announce Type: cross Abstract: Virtual contrast enhancement (VCE) synthesizes enhanced breast MR images from pre-contrast acquisitions. Modern latent generators offer strong image p

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

PRISMA: Improving the Accuracy-Latency Frontier of Diffusion-based PDE Solvers Using Physics-Informed Spectral Attention

DGX agent

arXiv:2512.01370v2 Announce Type: replace-cross Abstract: Diffusion-based solvers for partial differential equations (PDEs) are often bottle-necked by slow gradient-based test-time optimization routin

model-releasesarxiv-cs-ai
5 Aug 2026
Local Ai

Probing Character-level Transformers for the Spanish L-shaped Morphome

DGX agent

arXiv:2608.03452v1 Announce Type: new Abstract: When a transformer learns an irregular morphological pattern, what has it learned? Our test case is the Spanish L-shaped morphome, a complex irregular p

local-aiarxiv-cs-cl
5 Aug 2026
Model Releases

Provably Learning Multi-Head Attention with Queries

DGX agent

arXiv:2608.03294v1 Announce Type: new Abstract: We study the problem of learning multi-head softmax attention from black-box input-output access. The learner may query arbitrary real-valued token sequ

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

SciRet: A Compute-Aware Empirical Study of Retrieval and Reranking for Scientific RAG

DGX agent

arXiv:2608.03860v1 Announce Type: cross Abstract: We introduce SciRet, a compute-aware empirical study of retrieval-augmented generation for scientific question answering over CORD-19. Rather than pro

model-releasesarxiv-cs-ai
5 Aug 2026
Research

SeCo-SBIR: Semantically Consistent Prompt Learning for Zero-Shot Sketch-Based Image Retrieval

DGX agent

arXiv:2608.03120v1 Announce Type: new Abstract: Adapting CLIP for zero-shot sketch-based image retrieval (ZS-SBIR) via prompt learning faces a fundamental tension: the model must bridge the sketch-pho

researcharxiv-cs-cv
5 Aug 2026
Local Ai

SEER: A Self-Grounded Evidence Interface for Controlled Spatial Relation Classification

DGX agent

arXiv:2608.03631v1 Announce Type: new Abstract: Spatial relation questions require a model to identify the queried subject and object before comparing their layout. Yet a VLM can recognize both entiti

local-aiarxiv-cs-cv
5 Aug 2026
Model Releases

SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMs

DGX agent

arXiv:2608.03573v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL) exhibit fundamentally different behaviors in enhancing multi-task reasoning for large langu

model-releasesarxiv-cs-cl
5 Aug 2026
Local Ai

Stochastic Multiple Shooting Trajectory Optimization via Sequential Local Policy Evaluation

DGX agent

arXiv:2608.03978v1 Announce Type: new Abstract: Stochastic single shooting trajectory optimization methods such as Model Predictive Path Integral control (MPPI) have been widely adopted in robotics du

local-aiarxiv-cs-ro
5 Aug 2026
Model Releases

STREAM-VAE: Dual-Path Routing for Slow and Fast Dynamics in Vehicle Telemetry Anomaly Detection

DGX agent

arXiv:2511.15339v3 Announce Type: replace-cross Abstract: Automotive telemetry data exhibits slow drifts and fast spikes, often within the same sequence, making reliable anomaly detection challenging.

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Taming the Implicit: Dual-Channel Risk-Aware Reinforcement Fine-Tuning for Continual Multimodal Post-Training

DGX agent

arXiv:2608.03660v1 Announce Type: new Abstract: Reinforcement fine-tuning (RFT) is widely believed to inherently resist catastrophic forgetting in continual post-training of multimodal large language

model-releasesarxiv-cs-ai
5 Aug 2026
Research

Temporal Leakage in LLM Backtesting: Measurement, Validation, and Adjusted Scores

DGX agent

arXiv:2608.02985v1 Announce Type: cross Abstract: The standard check for contamination in LLM backtests is simple: compare scores before and after the training cutoff. We show this check is uninformat

researcharxiv-cs-cl
5 Aug 2026
Research

Test-Time Scaling in Reasoning LLMs: Inference Regimes, Evaluation, and Reproducibility

DGX agent

arXiv:2608.04001v1 Announce Type: cross Abstract: Large language models can solve substantially harder reasoning problems with more inference-time compute. The term 'test-time scaling,' however, now c

researcharxiv-cs-ai
5 Aug 2026
← Previous
1…660661662663664…1369
Next →