AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Model Releases

NetConfArena: An Executable Benchmark for LLM Agents in Closed-Loop Network Configuration

DGX agent

arXiv:2608.23179v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly attractive for automating network configuration, yet their reliability and failure patterns are poo

model-releasesarxiv-cs-ai
25 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Separating Voice from Age in COPD Screening

DGX agent

arXiv:2608.21599v1 Announce Type: cross Abstract: Voice has been proposed as a low-cost screening signal for chronic obstructive pulmonary disease (COPD). COPD is strongly age-associated and voice cha

model-releasesarxiv-cs-lg
25 Aug 2026
Model Releases

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding

DGX agent

arXiv:2607.14935v2 Announce Type: replace Abstract: Recent advances in video understanding have spanned motion, long video, and streaming interaction, driving this field toward real-world applications

model-releasesarxiv-cs-cv
25 Aug 2026
Model Releases

Wazobia Eval: A Benchmark for Nigerian Pidgin Emotion Understanding, Sarcasm Detection, and Cultural Reasoning

DGX agent

arXiv:2608.21369v1 Announce Type: cross Abstract: Nigerian Pidgin is one of Africa's most widely spoken languages, yet remains severely underrepresented in language model evaluation. Existing benchmar

model-releasesarxiv-cs-ai
25 Aug 2026
Model Releases

What do I want in AI? Here are just a few things. I want voice-enabled everything. I want things to be more proactive. I want to manage AI’s…

DGX agent

What do I want in AI? Here are just a few things. I want voice-enabled everything. I want things to be more proactive. I want to manage AI’s escalations (versus delegation). I want those escalations t

model-releasesallie-k--miller--x
25 Aug 2026
Model Releases

A Modular Agent for Reliable and Auditable Spatial Relation Verification in CT Scans

DGX agent

arXiv:2608.21140v1 Announce Type: cross Abstract: Reliable spatial understanding is an important prerequisite for future medical vision-language systems that aim to support radiological report generat

model-releasesarxiv-cs-ai
24 Aug 2026
Safety

aiXamine: Unified Black-Box Evaluation of Cross-Dimensional Trade-offs in LLM Safety, Security, and Privacy

DGX agent

arXiv:2608.20554v1 Announce Type: cross Abstract: The critical failure modes in deployed large language models (LLMs) are cross-dimensional: a model can score 99.3 in safety alignment while refusing o

safetyarxiv-cs-lg
24 Aug 2026
Model Releases

Automated Trajectory Evaluation for Mobile Agents via Step-Level Consequence Reasoning and Aggregation

DGX agent

arXiv:2608.20797v1 Announce Type: new Abstract: Evaluating language-guided mobile agents has recently shifted from rule-based to model-based approaches to achieve scalable and automated assessments. H

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

BC-Bench: Evaluating Agentic Engineering in a Domain-Specific Language for ERP

DGX agent

arXiv:2608.20851v1 Announce Type: cross Abstract: Agentic engineering systems have shown strong performance on general-purpose benchmarks, yet their effectiveness in enterprise resource planning (ERP)

model-releasesarxiv-cs-ai
24 Aug 2026
Local Ai

BF1: A Causal Dyadic Sparse-Attention Retrofit for Efficient Long-Context Transformers

DGX agent

arXiv:2608.20427v1 Announce Type: cross Abstract: Dense causal attention remains expensive at long context even when implemented with highly optimized exact kernels. We study BF1, a deterministic bloc

local-aiarxiv-cs-ai
24 Aug 2026
Tutorials

Difficulty-Calibrated Interpolation Paths for Conditional Flow Matching

DGX agent

arXiv:2608.21286v1 Announce Type: new Abstract: Conditional Flow Matching trains generative models by regressing a network onto the velocity of a prescribed noise-to-data interpolation path. The inter

tutorialsarxiv-cs-cv
24 Aug 2026
Model Releases

Evaluation-as-Search: Adaptive Discovery of Grounding Failures in Meeting Assistants

DGX agent

arXiv:2608.20392v1 Announce Type: cross Abstract: LLM-powered meeting assistants are deployed at scale, yet systematic evaluation of their grounding fidelity remains limited to static benchmarks that

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

FL-MAESTRO: Multi-Agent LLM Orchestration for Resource-Constrained Federated Learning

DGX agent

arXiv:2608.20518v1 Announce Type: new Abstract: In Federated Learning (FL), the communication topology is a runtime variable rather than a fixed design choice, since links and edge devices drop in and

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

Identity-Preserving Text-to-Video Generation via Agentic Enhancement and Semantic Repair

DGX agent

arXiv:2608.20749v1 Announce Type: new Abstract: Identity-preserving video generation aims to synthesize videos that follow natural-language instructions while maintaining the visual identity of a give

model-releasesarxiv-cs-cv
24 Aug 2026
Research

InverFill: One-Step Inversion for Enhanced Few-Step Diffusion Inpainting

DGX agent

arXiv:2603.23463v2 Announce Type: replace-cross Abstract: Recent diffusion-based models achieve photorealism in image inpainting but require many sampling steps, limiting practical use. Few-step text-

researcharxiv-cs-ai
24 Aug 2026
Model Releases

Is Vibe Coding Safe? Benchmarking Vulnerability of Agent-Generated Code in Real-World Tasks

DGX agent

arXiv:2512.03262v3 Announce Type: replace-cross Abstract: Vibe coding is a new software development paradigm in which human engineers prompt a large language model (LLM) agent to complete complex codi

model-releasesarxiv-cs-cl
24 Aug 2026
Research

Jokes Aside: Measuring the Semantic Distance of Double Meanings

DGX agent

arXiv:2608.21087v1 Announce Type: new Abstract: Large language models have significantly enriched the toolkit for computational humor research, particularly in the automated generation of jokes and pu

researcharxiv-cs-cl
24 Aug 2026
Applications

LoRC: Detecting AI-Generated Images via Low-Rank Collapse in Semantic Residuals

DGX agent

arXiv:2608.20882v1 Announce Type: new Abstract: Modern generators faithfully model macroscopic semantics, producing synthetic images that appear highly realistic. Consequently, decisive forensic cues

applicationsarxiv-cs-cv
24 Aug 2026
Research

Minimax Optimality of Score-Entropy Discrete Diffusion

DGX agent

arXiv:2608.20635v1 Announce Type: cross Abstract: Discrete diffusion models have demonstrated strong performance across a range of datasets, including natural language data and graph-structured data.

researcharxiv-cs-lg
24 Aug 2026
Applications

MotionPhys: Detecting AI-Generated Videos via Physical Consistency of Optical-Flow Trajectories

DGX agent

arXiv:2608.20770v1 Announce Type: new Abstract: Modern AI video generation models can produce videos with high visual fidelity and seemingly smooth temporal transitions. However, visual realism does n

applicationsarxiv-cs-cv
24 Aug 2026
Agents

Peer-Voted LLM-Agent Stress Tests Find Feed-Induced Lexical Convergence but No Reliable Matched-Exposure Advantage for Distributed Sources

DGX agent

arXiv:2608.20438v1 Announce Type: cross Abstract: Population-level behavior in large-language-model (LLM) agents cannot be characterized by single-agent benchmarks. We introduce PV-SST, a peer-voted s

agentsarxiv-cs-ai
24 Aug 2026
Research

Prediction certification cannot replace explanation certification: a competence envelope for trustworthy AI under compound stress

DGX agent

arXiv:2608.20825v1 Announce Type: new Abstract: Artificial intelligence systems increasingly make consequential judgments - which patient is deteriorating, which building is safe to enter, whether an

researcharxiv-cs-ai
24 Aug 2026
Model Releases

PSK at WMT 2026 MIST: Task-Specialized QLoRA Adapters for Multilingual Summarization and Question Answering

DGX agent

arXiv:2608.20757v1 Announce Type: cross Abstract: We describe the PSK submission to the WMT 2026 Multilingual Instruction Shared Task. Our system uses the 3.35B-parameter Tiny Aya Global model with th

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

RiskTraf: Risk-Extrapolated Residual Learning for Multi-Variate Traffic Flow Prediction

DGX agent

arXiv:2608.20656v1 Announce Type: cross Abstract: Traffic sensors commonly record flow, speed, and occupancy, but standard traffic flow forecasting benchmarks and models rarely exploit all three raw m

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

still more bad news for Anthropic. if they don’t scale back expectations for their IPO, it’s going to be a trainwreck.

DGX agent

still more bad news for Anthropic. if they don’t scale back expectations for their IPO, it’s going to be a trainwreck. NEW: Thomson Reuters wanted to rely less on Claude. So it built its own AI model

model-releasesgary-marcus--x
24 Aug 2026
Model Releases

Terminal Agents: A Survey of AI Agents in Command-Line Environments

DGX agent

arXiv:2608.20485v1 Announce Type: new Abstract: Large language model agents increasingly act through terminals, yet existing surveys disperse terminal-mediated behavior across software engineering, to

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

TH-GNN: Heterogeneous Temporal Graph Neural Networks for LLM-Agent Shilling Attack Detection

DGX agent

arXiv:2608.20376v1 Announce Type: new Abstract: LLM agents can now generate realistic shilling profiles, fluent reviews, and coherent ratings at scale, systematically defeating recommender-system defe

model-releasesarxiv-cs-cl
24 Aug 2026
Model Releases

This is what Qwen 3.8 27b is capable of

DGX agent

Try it here: https://ocean.blackbeardlabs.dev/ Model: Qwen 3.8 27b Q8_X_KL Unsloth Hardware: 3 x RTX3090 Harness: DeepSeek Harness Prompt: /goal I want you to create a **JavaScript + Node.js WebGL pro

model-releasesr-localllama
24 Aug 2026
Research

TracingFlow: A Simulation-Free Trajectory Inference Framework Based on Second-Order Dynamics

DGX agent

arXiv:2608.21070v1 Announce Type: cross Abstract: Inferring continuous system evolution from sparse temporal snapshots is a key challenge in generative modeling and single-cell omics. While Optimal Tr

researcharxiv-cs-ai
24 Aug 2026
Model Releases

VT-MUSE: Multimodal Unified Sequential Visuotactile Representation Learning for Manipulation

DGX agent

arXiv:2608.21290v1 Announce Type: cross Abstract: We propose VT-MUSE, a Multimodal Unified SEquential representation learning framework for visuotactilemanipulation. Existing approaches often encode v

model-releasesarxiv-cs-cv
24 Aug 2026
Research

When to Ponder: Adaptive Compute Allocation for Code Generation via Test-Time Training

DGX agent

arXiv:2601.00894v2 Announce Type: replace-cross Abstract: Large language models apply uniform computation to all inputs, regardless of difficulty. We propose PonderTTT, a gating strategy using the TTT

researcharxiv-cs-cl
24 Aug 2026
Model Releases

When Words Are Safe But Actions Kill: Probing Physical Jailbreak Beyond Textual Jailbreak in Hidden-State Risk Space

DGX agent

arXiv:2607.15218v2 Announce Type: replace Abstract: Large language models (LLMs) increasingly serve as high-level planners for embodied agents, where linguistically benign instructions can become unsa

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

28 TPS on Qwen2.5-7B across two separate cloud regions over public WAN using speculative decoding + CUDA Graphs [P]

DGX agent

been building ShardFlow for the past few months, a distributed LLM inference framework that splits any HuggingFace transformer across N GPU machines and uses neural speculative decoding to deal with W

model-releasesr-machinelearning
23 Aug 2026
Model Releases

Fixed the MTP head on Ornith1.5 35B A3B. +3% TPS -33% wall clock

DGX agent

I love the Ornith 35B local models, 1.0 has been running my HAM radio rig for me. I have a hackRF receiver and a 5 watt quansheng portable the both run headless through the PC. I tried out the new Orn

model-releasesr-localllama
22 Aug 2026
Model Releases

Great paper if you are tracking progress in recursive self-improvement (RSI). (bookmark it) There is so much hype around RSI, so I think it'…

DGX agent

Great paper if you are tracking progress in recursive self-improvement (RSI). (bookmark it) There is so much hype around RSI, so I think it's worth understanding why current models are not able to do

model-releasesdair-ai--x
22 Aug 2026
Model Releases

Qwen 3.8 27b - PI AGENT vs OPENCODE - another smaple

DGX agent

That is the second comparison and the last one. I will not be spamming again ;) Continuation from: https://www.reddit.com/r/LocalLLaMA/comments/1vu0u2v/qwen_38_27b_pi_agent_vs_opencode/ That is one of

model-releasesr-localllama
22 Aug 2026
Agents

This is one of the most effective ways to improve your agentic workflows. If you are building computer-use agents, this one is worth your ti…

DGX agent

This is one of the most effective ways to improve your agentic workflows. If you are building computer-use agents, this one is worth your time. Task Model Induction takes a raw recording of someone wo

agentsdair-ai--x
22 Aug 2026
Local Ai

What if Chain-of-Thought wasn’t lossy? Exploring reversible logic (Toffoli/Fredkin-style) for edge LLMs

DGX agent

Right now standard CoT is a one-way street. You generate forward, dump a pile of scratchpad tokens into the KV cache, and pray the model doesn’t hallucinate halfway through. On phones/laptops that cre

local-air-localllama
22 Aug 2026
Model Releases

Automatic bioinformatic software named entity recognition from literature

DGX agent

arXiv:2608.19201v1 Announce Type: cross Abstract: Bioinformatics software and databases are essential components of modern life science research, yet their mentions in the scientific literature are of

model-releasesarxiv-cs-ai
21 Aug 2026
Model Releases

ContractScrub: A benchmark for final review of legal contracts

DGX agent

arXiv:2608.20204v1 Announce Type: new Abstract: Legal work, with its heavy reliance on processing large amounts of text, is often considered one of the domains most exposed to the use of LLMs. Contrac

model-releasesarxiv-cs-ai
21 Aug 2026
Hardware

FlashPrefill V2: Block-Sparse Prefill Attention for Long-Context LLM Serving

DGX agent

arXiv:2608.19758v1 Announce Type: new Abstract: Long-context modeling is a pivotal capability for Large Language Models, yet the quadratic complexity of attention remains a critical bottleneck, partic

hardwarearxiv-cs-cl
21 Aug 2026
Model Releases

FleetSieve: Decision-Critical Profiling for SLO-Aware LLM Fleet Configuration

DGX agent

arXiv:2608.19659v1 Announce Type: new Abstract: Choosing tensor-parallel (TP) degrees and replica counts for an LLM serving fleet is difficult because performance is not monotonic in TP and the feasib

model-releasesarxiv-cs-lg
21 Aug 2026
Safety

Flow Matching-Based PET Image Reconstruction

DGX agent

arXiv:2608.20112v1 Announce Type: cross Abstract: Generative models have shown strong potential for positron emission tomography (PET) image reconstruction. Although diffusion model-based reconstructi

safetyarxiv-cs-cv
21 Aug 2026
Model Releases

From Retrieved Context to Runtime Control: Adaptive Compression for Edge-based RAG

DGX agent

arXiv:2608.19535v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves language-model responses by grounding generation in external passages, which comes with overhead: retrieve

model-releasesarxiv-cs-ai
21 Aug 2026
Model Releases

GRACE: Grounded Reasoning via Adapter Composition and Evidence-Aware Calibration for Educational Visual Question Answering

DGX agent

arXiv:2608.19355v1 Announce Type: cross Abstract: Educational visual question answering, or VQA, requires models to solve curriculum-oriented multiple-choice questions using both language and visual e

model-releasesarxiv-cs-cv
21 Aug 2026
Model Releases

K3 Usage Comparisons?

DGX agent

I'm confused by something I've noted in the recent discourse since Kimi K3 has been released on subscription. Has anyone done the (I know Ollama should display this) math as well on how much K3 is usi

model-releasesr-ollama
21 Aug 2026
Safety

Learning When to Think: Adaptive Reasoning for Test-Time Compute Allocation

DGX agent

arXiv:2608.20256v1 Announce Type: new Abstract: Reasoning language models trained with reinforcement learning typically operate under a fixed token budget rather than an explicitly adaptive one, which

safetyarxiv-cs-ai
21 Aug 2026
Research

Linguistic Holonomy and Statistical Watermarks: Inner Geometry of Meaning-Preserving Transformations

DGX agent

arXiv:2608.19369v1 Announce Type: new Abstract: Statistical watermarks for language models live in the freedom of the signifier: they choose among tokens that are nearly equivalent in meaning, and the

researcharxiv-cs-cl
21 Aug 2026
← Previous
1…430431432433434…1371
Next →