AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Research

Wired for Overconfidence: A Mechanistic Perspective on Inflated Verbalized Confidence in LLMs

DGX agent

arXiv:2604.01457v3 Announce Type: replace Abstract: Large language models are often not just wrong, but confidently wrong: when they produce factually incorrect answers, they tend to verbalize overly

researcharxiv-cs-cl
28 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Breaking the Data Barrier in Learning Symbolic Computation: A Case Study on Variable Ordering Suggestion for Cylindrical Algebraic Decomposition

DGX agent

arXiv:2601.13731v2 Announce Type: replace-cross Abstract: Symbolic computation, powered by modern computer algebra systems, has important applications in mathematical reasoning through exact deep comp

model-releasesarxiv-cs-lg
27 Jul 2026
Safety

Color Me Correctly: Bridging Perceptual Color Spaces and Text Embeddings for Improved Diffusion Generation

DGX agent

arXiv:2509.10058v2 Announce Type: replace Abstract: Accurate color alignment in text-to-image (T2I) generation is critical for applications such as fashion, product visualization, and interior design,

safetyarxiv-cs-cv
27 Jul 2026
Research

Complexity Bounds and Approaches to Learning Projected Gradient Descent Solver Iterates

DGX agent

arXiv:2607.22467v1 Announce Type: new Abstract: Data scarcity poses a fundamental challenge in training generative models to produce initial guesses for parametric optimization problems that are other

researcharxiv-cs-lg
27 Jul 2026
Safety

Constraint-Driven Synthesis of Hyper Petri Nets

DGX agent

arXiv:2607.22062v1 Announce Type: cross Abstract: This paper addresses the modeling and synthesis of constrained robotic system behaviors using Petri nets (PNs). It investigates how to construct model

safetyarxiv-cs-ro
27 Jul 2026
Local Ai

FBLayout: Optimizing Memory Layout for Efficient LLM Finetuning on Mobile GPUs

DGX agent

arXiv:2607.21624v1 Announce Type: cross Abstract: Transformer-based models have enabled unprecedented capabilities across language, vision, and multimodal tasks. On-device fine-tuning of transformer m

local-aiarxiv-cs-lg
27 Jul 2026
Hardware

Good move by @JensenHuang. The Nvidia letter is well written and worth reading. As we saw with the OpenAI-Hugging Face hack, we need open mo…

DGX agent

Good move by @JensenHuang. The Nvidia letter is well written and worth reading. As we saw with the OpenAI-Hugging Face hack, we need open models and harnesses for defense. Lets stop believing the PR t

hardwareandrew-ng--x
27 Jul 2026
Local Ai

I want to run Kimi K3 at home, so I’m trying to make 2.8T-scale experimentation cheaper

DGX agent

Hey r/LocalLLaMA, I’m a retired engineer with a background in distributed computing, currently running a 1-person startup. Like many people here, I’d love to experiment with 2T+ MoE models locally. Th

local-air-localllama
27 Jul 2026
Model Releases

InteractComp: Evaluating Search Agents With Ambiguous Queries

DGX agent

arXiv:2510.24668v2 Announce Type: replace Abstract: Language agents have demonstrated remarkable potential in web search and information retrieval. However, many search-agent benchmarks assume that us

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Kat Coder 2.5 is insane. Especially considering I ran it at Q4_K_M

DGX agent

I tested Kat Coder 2.5 with this prompt: Create a spaceship game inspired by Star Fox using vanilla Three.js and HTML. It should have at least five levels, keyboard and mouse controls, enemies, and a

model-releasesr-localllama
27 Jul 2026
Model Releases

moonshotai/Kimi-K3

DGX agent

moonshotai/Kimi-K3 As promised earlier this month, Moonshot have released the weights for their excellent 2.8 trillion parameter Kimi K3. They're a hefty 1.56TB on Hugging Face. Kimi introduced their

model-releasessimon-willison
27 Jul 2026
Hardware

NVIDIA Ising Enables Fully Automated Quantum Computer Calibration with Enhanced In-Context Learning

DGX agent

NVIDIA Ising Calibration 1.5 is a 31‑billion‑parameter vision‑language model that diagnoses and tunes quantum processors, delivering state‑of‑the‑art zero‑shot and in‑context learning on the QCalEval

hardwarenvidia-developer
27 Jul 2026
Model Releases

OpenNavMap: Multi-Session Appearance-Based Topometric Mapping for Scalable Visual Navigation

DGX agent

arXiv:2601.12291v2 Announce Type: replace-cross Abstract: Scalable and maintainable maps are fundamental to large-scale navigation and the long-term deployment of robots in real-world environments. Ho

model-releasesarxiv-cs-cv
27 Jul 2026
Hardware

Proud to be an inaugural partner alongside @NVIDIA and other industry leaders in the Open Secure AI Alliance. We look forward to acceleratin…

DGX agent

Proud to be an inaugural partner alongside @NVIDIA and other industry leaders in the Open Secure AI Alliance. We look forward to accelerating the adoption and trust of open models and harnesses throug

hardwareharrison-chase--x
27 Jul 2026
Model Releases

QC-PHAST Search: Classical--Quantum Query Benchmarks for Finite-Pool Rare-Regime Discovery

DGX agent

arXiv:2607.21995v1 Announce Type: cross Abstract: Rare-regime discovery in parameterized dynamical systems is an active-search problem: find one verified parameter at which a scientifically defined qu

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills

DGX agent

arXiv:2607.22529v1 Announce Type: new Abstract: LLM training is shifting from manual design and annotation to interaction-driven self-evolution. However, existing self-evolutionary methods face a fund

model-releasesarxiv-cs-cl
27 Jul 2026
Applications

Start owning your own intelligence today with Kimi K3 on Fireworks. Serve it, fine-tune it, and put it into production. What are you waiting…

DGX agent

**Summary:** Kimi K3 is a publicly available 3‑trillion‑parameter language model now live on Fireworks, offering 1 million-token context windows with native vision and reasoning that rivals top closed

applicationsfireworks-ai--x
27 Jul 2026
Tutorials

Statistical mechanics of extensive-width Bayesian neural networks near interpolation

DGX agent

arXiv:2505.24849v2 Announce Type: replace-cross Abstract: For three decades statistical mechanics has been providing a framework to analyse neural networks. However, the theoretically tractable models

tutorialsarxiv-cs-lg
27 Jul 2026
Model Releases

Stop to Decide: Latency-Aware Proprioceptive Navigation Primitives for Mapping-Free Quadruped Inspection

DGX agent

arXiv:2607.11204v2 Announce Type: replace Abstract: Onboard quadruped inspection systems often share limited compute between perception and navigation, reducing the rate at which event-triggered contr

model-releasesarxiv-cs-ro
27 Jul 2026
Model Releases

Time-Reversed Imaging: A Multimodal Benchmark and Framework for Reconstructing Past Human-Environment Interactions

DGX agent

arXiv:2607.22352v1 Announce Type: new Abstract: We introduce time-reversed imaging, a new paradigm that infers what just happened in a scene from fading multimodal traces. Instead of extrapolating or

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Toward User-Conditioned Evaluation of Personal LLM Agents under Temporal Interventions

DGX agent

arXiv:2607.21635v1 Announce Type: new Abstract: Personal agents maintain memories, learned skills, tool configurations, and policy state that evolve with each user. Existing agent benchmarks often eva

model-releasesarxiv-cs-lg
27 Jul 2026
Safety

TRACE-ROUTER: Task-Consistent and Adaptive Online Routing for Agentic AI

DGX agent

arXiv:2607.22465v1 Announce Type: cross Abstract: Routing to select large language models (LLMs) with different cost-quality trade-offs has become a fundamental deployment feature of enterprise AI. Ex

safetyarxiv-cs-lg
27 Jul 2026
Agents

Want to go deeper? Join Moonshot AI and Together AI for a technical webinar on how K3 was built and how to use it for production agent workf…

DGX agent

Together AI has released the Kimi K3 model on its platform as a Day‑0 launch partner for Moonshot AI’s open frontier agentic model, which supports long‑running workflows across code, tools, vision and

agentstogether-ai--x
27 Jul 2026
Model Releases

BeeLlama.cpp v0.4.1: KVarN, KV precision tail, q2_0-q3_1 KV cache, improved support. KLD benchmarks: tail 1024 makes kvarn5 and q6_0 match q8_0, for much less VRAM

DGX agent

TL;DR llama.cpp fork with more KV cache quantization features, with all claims supported by benchmarks: KVarN, KV cache precision tail, additional types of standard KV cache (q2_0-q3_1, q6_0, q6_1), a

model-releasesr-localllama
26 Jul 2026
Local Ai

[Paper] RecGPT-V3 Technical Report

DGX agent

Large language models (LLMs) are transforming recommender systems from matching co-occurrence patterns in historical behavior toward reasoning about the intent that drives it. RecGPT-V1 pioneered this

local-air-localllama
26 Jul 2026
Model Releases

Help me complete my AI collection

DGX agent

I’m building the ultimate AI tool vault, but every great collection has a few missing pieces. Note: I will react to every comment AI's currently installed: Qwen3.5-0.8B-UD-Q4_K_XL.gguf(classification)

model-releasesr-localllama
25 Jul 2026
Model Releases

Llama.cpp now has full MCP support!

DGX agent

After a long and grueling effort spearheaded by ngxson, llama.cpp now fully supports MCP for all protocols. Over-the-web HTTP servers were already supported in the client (since they don't require any

model-releasesr-localllama
25 Jul 2026
Model Releases

MI50 power curve tests

DGX agent

tests done power limiting the GPU on LACT - real power usage varies wildy at 20W it ranges from 25W to 56W same behavior happens on every setting prompt for the test runs: https://github.com/lukesdevl

model-releasesr-localllama
25 Jul 2026
Model Releases

Ollama Qwen3.6:35b randomly stops outputting tokens

DGX agent

RTX 4070, 32gb system ram, Linux. NVIDIA-SMI 610.43.03, KMD Version: 610.43.03, CUDA UMD Version: 13.3 Systemd service modifications: [Service] Environment='OLLAMA_HOST=0.0.0.0:11434' Environment='OLL

model-releasesr-localllama
25 Jul 2026
Model Releases

A Comparative Evaluation of Embeddings and LLMs in a Greek Book Publisher Setting - The CUP Dataset

DGX agent

arXiv:2607.21274v1 Announce Type: cross Abstract: We present CUP, a Greek book retrieval benchmark consisting of 868 catalog records and 104 expert-annotated queries with graded relevance judgments. W

model-releasesarxiv-cs-ai
24 Jul 2026
Safety

Adaptive Depth Sparse Framework: Similarity-Driven Resource Allocation for Pre-Trained LLMs

DGX agent

arXiv:2607.21291v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong generation and reasoning performance, but the Transformer architecture incurs high inference cost. Existing

safetyarxiv-cs-cl
24 Jul 2026
Agents

Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment

DGX agent

arXiv:2607.21437v1 Announce Type: new Abstract: Deep learning models can effectively use Rapid Evaporative Ionization Mass Spectrometry (REIMS) data for surgical margin assessment. However, their clin

agentsarxiv-cs-ai
24 Jul 2026
Model Releases

Agentic Designer: Progressive Multi-Agent Collaboration for Structure-Aware Interior Layout Generation

DGX agent

arXiv:2607.20866v1 Announce Type: new Abstract: Generating realistic interior furniture layouts that strictly adhere to architectural constraints (e.g., walls, doors, and windows) remains a fundamenta

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

An interesting thing I'm observing from the blog/system card is that on a good chunk of the reported benchmarks (~20-30% from a skim), Opus …

DGX agent

An interesting thing I'm observing from the blog/system card is that on a good chunk of the reported benchmarks (~20-30% from a skim), Opus 5 max thinking leads to a degradation in performance compare

model-releasesjerry-liu--x
24 Jul 2026
Research

Bridging the Gap Between Plausibility and Admissibility: Constraint-Aware Flow Maps for Dynamic Graph Systems

DGX agent

arXiv:2607.21421v1 Announce Type: new Abstract: Generative models can support decision-making under uncertainty by producing ensembles of plausible future system trajectories, but statistical plausibi

researcharxiv-cs-ai
24 Jul 2026
Safety

Confidently Deceptive: How Confidence Amplifies the Risk of LLM Deception

DGX agent

arXiv:2607.20444v1 Announce Type: cross Abstract: Large language models (LLMs) can produce deceptive responses: outputs that mislead users in service of a contextually or experimentally induced goal.

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

Deblurring in the Wild: A Real-World Image Deblurring Dataset from Smartphone High-Speed Videos

DGX agent

arXiv:2506.19445v4 Announce Type: cross Abstract: We introduce the largest real-world image deblurring dataset constructed from smartphone slow-motion videos. Using 240 frames captured over one second

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Directional Hallucinations: Ideological Drift in News-Grounded LLM Question Answering

DGX agent

arXiv:2607.20487v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to answer questions about political information, including in election-adjacent information settings

researcharxiv-cs-ai
24 Jul 2026
Tutorials

ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing

DGX agent

arXiv:2607.21529v1 Announce Type: cross Abstract: Test-Time Tuning (TTT) on pretrained diffusion models has emerged as a powerful paradigm for video editing. However, there exists a foundational misma

tutorialsarxiv-cs-ai
24 Jul 2026
Research

Fitting Generalized Power Diagrams to 3D Image Data: A Prerequisite for Virtual Materials Testing

DGX agent

arXiv:2507.14268v2 Announce Type: replace Abstract: This paper reviews algorithmic and modeling approaches for fitting generalized power diagrams to three-dimensional image data, a key step in virtual

researcharxiv-cs-cv
24 Jul 2026
Model Releases

From Attention to Frequency: Integration of Vision Transformer and FFT-ReLU for Enhanced Image Deblurring

DGX agent

arXiv:2511.10806v1 Announce Type: cross Abstract: Image deblurring is vital in computer vision, aiming to recover sharp images from blurry ones caused by motion or camera shake. While deep learning ap

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

How Many Bits Can an Adapter Write? Measuring the Capacity and Memorization of Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2607.21351v1 Announce Type: new Abstract: A LoRA adapter is a few megabytes that almost everyone treats as a skill rather than a record of the data behind it. We put that assumption on a scale.

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

I built an open-source multi-agent SDLC harness that beats a cold Claude Code run on large repos, by learning the repo once. Real benchmarks (incl. where it loses) inside. [P]

DGX agent

Built an open-source AI coding agent that was 7%–75% cheaper than a cold 'claude -p' run on 6/6 well-localized tasks across repositories up to ~82k LOC. The biggest difference: Cold agent: 6.83, 207 t

model-releasesr-machinelearning
24 Jul 2026
Model Releases

Instruct-FD: Can Your Full-Duplex Speech System Follow Turn-Taking Instructions?

DGX agent

arXiv:2607.20460v1 Announce Type: cross Abstract: Current full-duplex (FD) spoken dialogue systems can produce fluid interactions, yet it remains unclear whether they can adapt their turn-taking behav

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Is Deep Research Reliable? Misleading Knowledge Induces False Conclusions

DGX agent

arXiv:2607.20891v1 Announce Type: new Abstract: Deep Research agents extend LLM-based assistants into long-horizon workflows involving planning, retrieval, evidence synthesis, and report generation, y

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Just over 6 months later, Opus 5 now produces near-superhuman level spreadsheets and slide decks that match what a consultant would make. Th…

DGX agent

Just over 6 months later, Opus 5 now produces near-superhuman level spreadsheets and slide decks that match what a consultant would make. Things are changing fast. Media I'm hearing from many folks ac

model-releasesboris-cherny--x
24 Jul 2026
Applications

Learning to Detect UI Principle Violations via Reinforcement Learning

DGX agent

arXiv:2607.20690v1 Announce Type: new Abstract: Small language models and coding agents increasingly generate web front-end code, yet their outputs are typically evaluated primarily for functional cor

applicationsarxiv-cs-cl
24 Jul 2026
Model Releases

Learning to Navigate Efficiently with Only 0.58M Trainable Parameters

DGX agent

arXiv:2607.11029v2 Announce Type: replace-cross Abstract: Recent progress in visual navigation has largely been driven by scale: end-to-end policies with hundreds of millions of parameters trained on

model-releasesarxiv-cs-cv
24 Jul 2026
← Previous
1…584585586587588…1371
Next →