AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
92,405Total entries
1Added by human
92,404Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,927 results
Tutorials

Statistical mechanics of extensive-width Bayesian neural networks near interpolation

DGX agent

arXiv:2505.24849v2 Announce Type: replace-cross Abstract: For three decades statistical mechanics has been providing a framework to analyse neural networks. However, the theoretically tractable models

tutorialsarxiv-cs-lg
27 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Stop to Decide: Latency-Aware Proprioceptive Navigation Primitives for Mapping-Free Quadruped Inspection

DGX agent

arXiv:2607.11204v2 Announce Type: replace Abstract: Onboard quadruped inspection systems often share limited compute between perception and navigation, reducing the rate at which event-triggered contr

model-releasesarxiv-cs-ro
27 Jul 2026
Model Releases

Time-Reversed Imaging: A Multimodal Benchmark and Framework for Reconstructing Past Human-Environment Interactions

DGX agent

arXiv:2607.22352v1 Announce Type: new Abstract: We introduce time-reversed imaging, a new paradigm that infers what just happened in a scene from fading multimodal traces. Instead of extrapolating or

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Toward User-Conditioned Evaluation of Personal LLM Agents under Temporal Interventions

DGX agent

arXiv:2607.21635v1 Announce Type: new Abstract: Personal agents maintain memories, learned skills, tool configurations, and policy state that evolve with each user. Existing agent benchmarks often eva

model-releasesarxiv-cs-lg
27 Jul 2026
Safety

TRACE-ROUTER: Task-Consistent and Adaptive Online Routing for Agentic AI

DGX agent

arXiv:2607.22465v1 Announce Type: cross Abstract: Routing to select large language models (LLMs) with different cost-quality trade-offs has become a fundamental deployment feature of enterprise AI. Ex

safetyarxiv-cs-lg
27 Jul 2026
Agents

Want to go deeper? Join Moonshot AI and Together AI for a technical webinar on how K3 was built and how to use it for production agent workf…

DGX agent

Together AI has released the Kimi K3 model on its platform as a Day‑0 launch partner for Moonshot AI’s open frontier agentic model, which supports long‑running workflows across code, tools, vision and

agentstogether-ai--x
27 Jul 2026
Model Releases

BeeLlama.cpp v0.4.1: KVarN, KV precision tail, q2_0-q3_1 KV cache, improved support. KLD benchmarks: tail 1024 makes kvarn5 and q6_0 match q8_0, for much less VRAM

DGX agent

TL;DR llama.cpp fork with more KV cache quantization features, with all claims supported by benchmarks: KVarN, KV cache precision tail, additional types of standard KV cache (q2_0-q3_1, q6_0, q6_1), a

model-releasesr-localllama
26 Jul 2026
Local Ai

[Paper] RecGPT-V3 Technical Report

DGX agent

Large language models (LLMs) are transforming recommender systems from matching co-occurrence patterns in historical behavior toward reasoning about the intent that drives it. RecGPT-V1 pioneered this

local-air-localllama
26 Jul 2026
Model Releases

Help me complete my AI collection

DGX agent

I’m building the ultimate AI tool vault, but every great collection has a few missing pieces. Note: I will react to every comment AI's currently installed: Qwen3.5-0.8B-UD-Q4_K_XL.gguf(classification)

model-releasesr-localllama
25 Jul 2026
Model Releases

Llama.cpp now has full MCP support!

DGX agent

After a long and grueling effort spearheaded by ngxson, llama.cpp now fully supports MCP for all protocols. Over-the-web HTTP servers were already supported in the client (since they don't require any

model-releasesr-localllama
25 Jul 2026
Model Releases

MI50 power curve tests

DGX agent

tests done power limiting the GPU on LACT - real power usage varies wildy at 20W it ranges from 25W to 56W same behavior happens on every setting prompt for the test runs: https://github.com/lukesdevl

model-releasesr-localllama
25 Jul 2026
Model Releases

Ollama Qwen3.6:35b randomly stops outputting tokens

DGX agent

RTX 4070, 32gb system ram, Linux. NVIDIA-SMI 610.43.03, KMD Version: 610.43.03, CUDA UMD Version: 13.3 Systemd service modifications: [Service] Environment='OLLAMA_HOST=0.0.0.0:11434' Environment='OLL

model-releasesr-localllama
25 Jul 2026
Model Releases

A Comparative Evaluation of Embeddings and LLMs in a Greek Book Publisher Setting - The CUP Dataset

DGX agent

arXiv:2607.21274v1 Announce Type: cross Abstract: We present CUP, a Greek book retrieval benchmark consisting of 868 catalog records and 104 expert-annotated queries with graded relevance judgments. W

model-releasesarxiv-cs-ai
24 Jul 2026
Safety

Adaptive Depth Sparse Framework: Similarity-Driven Resource Allocation for Pre-Trained LLMs

DGX agent

arXiv:2607.21291v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong generation and reasoning performance, but the Transformer architecture incurs high inference cost. Existing

safetyarxiv-cs-cl
24 Jul 2026
Agents

Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment

DGX agent

arXiv:2607.21437v1 Announce Type: new Abstract: Deep learning models can effectively use Rapid Evaporative Ionization Mass Spectrometry (REIMS) data for surgical margin assessment. However, their clin

agentsarxiv-cs-ai
24 Jul 2026
Model Releases

Agentic Designer: Progressive Multi-Agent Collaboration for Structure-Aware Interior Layout Generation

DGX agent

arXiv:2607.20866v1 Announce Type: new Abstract: Generating realistic interior furniture layouts that strictly adhere to architectural constraints (e.g., walls, doors, and windows) remains a fundamenta

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

An interesting thing I'm observing from the blog/system card is that on a good chunk of the reported benchmarks (~20-30% from a skim), Opus …

DGX agent

An interesting thing I'm observing from the blog/system card is that on a good chunk of the reported benchmarks (~20-30% from a skim), Opus 5 max thinking leads to a degradation in performance compare

model-releasesjerry-liu--x
24 Jul 2026
Research

Bridging the Gap Between Plausibility and Admissibility: Constraint-Aware Flow Maps for Dynamic Graph Systems

DGX agent

arXiv:2607.21421v1 Announce Type: new Abstract: Generative models can support decision-making under uncertainty by producing ensembles of plausible future system trajectories, but statistical plausibi

researcharxiv-cs-ai
24 Jul 2026
Safety

Confidently Deceptive: How Confidence Amplifies the Risk of LLM Deception

DGX agent

arXiv:2607.20444v1 Announce Type: cross Abstract: Large language models (LLMs) can produce deceptive responses: outputs that mislead users in service of a contextually or experimentally induced goal.

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

Deblurring in the Wild: A Real-World Image Deblurring Dataset from Smartphone High-Speed Videos

DGX agent

arXiv:2506.19445v4 Announce Type: cross Abstract: We introduce the largest real-world image deblurring dataset constructed from smartphone slow-motion videos. Using 240 frames captured over one second

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Directional Hallucinations: Ideological Drift in News-Grounded LLM Question Answering

DGX agent

arXiv:2607.20487v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to answer questions about political information, including in election-adjacent information settings

researcharxiv-cs-ai
24 Jul 2026
Tutorials

ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing

DGX agent

arXiv:2607.21529v1 Announce Type: cross Abstract: Test-Time Tuning (TTT) on pretrained diffusion models has emerged as a powerful paradigm for video editing. However, there exists a foundational misma

tutorialsarxiv-cs-ai
24 Jul 2026
Research

Fitting Generalized Power Diagrams to 3D Image Data: A Prerequisite for Virtual Materials Testing

DGX agent

arXiv:2507.14268v2 Announce Type: replace Abstract: This paper reviews algorithmic and modeling approaches for fitting generalized power diagrams to three-dimensional image data, a key step in virtual

researcharxiv-cs-cv
24 Jul 2026
Model Releases

From Attention to Frequency: Integration of Vision Transformer and FFT-ReLU for Enhanced Image Deblurring

DGX agent

arXiv:2511.10806v1 Announce Type: cross Abstract: Image deblurring is vital in computer vision, aiming to recover sharp images from blurry ones caused by motion or camera shake. While deep learning ap

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

How Many Bits Can an Adapter Write? Measuring the Capacity and Memorization of Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2607.21351v1 Announce Type: new Abstract: A LoRA adapter is a few megabytes that almost everyone treats as a skill rather than a record of the data behind it. We put that assumption on a scale.

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

I built an open-source multi-agent SDLC harness that beats a cold Claude Code run on large repos, by learning the repo once. Real benchmarks (incl. where it loses) inside. [P]

DGX agent

Built an open-source AI coding agent that was 7%–75% cheaper than a cold 'claude -p' run on 6/6 well-localized tasks across repositories up to ~82k LOC. The biggest difference: Cold agent: 6.83, 207 t

model-releasesr-machinelearning
24 Jul 2026
Model Releases

Instruct-FD: Can Your Full-Duplex Speech System Follow Turn-Taking Instructions?

DGX agent

arXiv:2607.20460v1 Announce Type: cross Abstract: Current full-duplex (FD) spoken dialogue systems can produce fluid interactions, yet it remains unclear whether they can adapt their turn-taking behav

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Is Deep Research Reliable? Misleading Knowledge Induces False Conclusions

DGX agent

arXiv:2607.20891v1 Announce Type: new Abstract: Deep Research agents extend LLM-based assistants into long-horizon workflows involving planning, retrieval, evidence synthesis, and report generation, y

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Just over 6 months later, Opus 5 now produces near-superhuman level spreadsheets and slide decks that match what a consultant would make. Th…

DGX agent

Just over 6 months later, Opus 5 now produces near-superhuman level spreadsheets and slide decks that match what a consultant would make. Things are changing fast. Media I'm hearing from many folks ac

model-releasesboris-cherny--x
24 Jul 2026
Applications

Learning to Detect UI Principle Violations via Reinforcement Learning

DGX agent

arXiv:2607.20690v1 Announce Type: new Abstract: Small language models and coding agents increasingly generate web front-end code, yet their outputs are typically evaluated primarily for functional cor

applicationsarxiv-cs-cl
24 Jul 2026
Model Releases

Learning to Navigate Efficiently with Only 0.58M Trainable Parameters

DGX agent

arXiv:2607.11029v2 Announce Type: replace-cross Abstract: Recent progress in visual navigation has largely been driven by scale: end-to-end policies with hundreds of millions of parameters trained on

model-releasesarxiv-cs-cv
24 Jul 2026
Research

Logic Programming Semantics for Causal Processes

DGX agent

arXiv:2607.21233v1 Announce Type: new Abstract: Motivated by challenging modelling issues in the life sciences, we investigate the relationship between logic programming semantics and the eventual sta

researcharxiv-cs-ai
24 Jul 2026
Model Releases

MVEI & EmObserver: Empowering MLLM-Oriented Visual Emotional Intelligence via Emotion Statement Judgement

DGX agent

arXiv:2607.21061v1 Announce Type: new Abstract: Affective Image Content Analysis (AICA) aims to recognize and understand emotions elicited by visual content, representing an indispensable step toward

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

New research with Microsoft's and colleagues on training agents inside the harnesses they actually run in. (bookmark it) Why it matters: Age…

DGX agent

New research with Microsoft's and colleagues on training agents inside the harnesses they actually run in. (bookmark it) Why it matters: Agents today live inside elaborate harnesses like Claude Code,

model-releasesdair-ai--x
24 Jul 2026
Model Releases

OpenForgeRL: Train Harness-native Agents in Any Environment

DGX agent

arXiv:2607.21557v1 Announce Type: new Abstract: Modern AI agents rely on elaborate inference harnesses such as Claude Code, Codex, and OpenClaw to drive multi-turn reasoning, tool use, and access to e

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Position Bias is Hidden Behind Ceiling Effects: A Permutation Diagnostic for LLM Benchmarks

DGX agent

arXiv:2607.20864v1 Announce Type: cross Abstract: Position bias in multiple-choice LLM evaluation is widely cited as a confound in capability comparisons, but published measurements rely on single ans

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

Preference Tuning as Spectral Update Reorganization

DGX agent

arXiv:2607.20438v1 Announce Type: cross Abstract: Preference-based post-training is usually understood through endpoint behavior, yet the learned update that produces this behavior remains largely opa

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

RUMBA: Russian User Memory Benchmark

DGX agent

arXiv:2607.21447v1 Announce Type: cross Abstract: The ability to handle long-term memory in LLMs is becoming increasingly critical, yet existing benchmarks remain English-centric and rely on aggregate

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

SenCos-GEM: SENet-Calibrated and Law-of-Cosines-Constrained Geometry-Enhanced Molecular Representation for Property Prediction

DGX agent

arXiv:2607.20551v1 Announce Type: cross Abstract: Effective molecular representation learning is crucial for accurate molecular property prediction. Recently, numerous self-supervised learning (SSL) a

model-releasesarxiv-cs-ai
24 Jul 2026
Applications

Smooth Neural Point Processes via B-Splines

DGX agent

arXiv:2607.21098v1 Announce Type: new Abstract: Temporal point processes (TPPs) provide a general and flexible framework for modeling sequences of events in continuous time. Neural networks have been

applicationsarxiv-cs-lg
24 Jul 2026
Model Releases

Spectral-Spatial Synergistic Guided Network for Hyperspectral Salient Object Detection

DGX agent

arXiv:2607.21032v1 Announce Type: new Abstract: Hyperspectral salient object detection aims to identify visually salient regions from hyperspectral images. Existing methods often fail because they fun

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

T-STAR: A Large-Scale Benchmark for Spatio-Temporal Panoptic Scene Graph Generation in Satellite Video

DGX agent

arXiv:2607.21228v1 Announce Type: new Abstract: Structured understanding of satellite video is essential for advancing dynamic geospatial scene analysis from low-level perception to high-level cogniti

model-releasesarxiv-cs-cv
24 Jul 2026
Safety

V-DEAL: Diagnosing Video Safety De-Calibration as an Understanding-Refusal Coupling Failure

DGX agent

arXiv:2607.21151v1 Announce Type: new Abstract: As Video Large Language Models are increasingly deployed in real-world applications, ensuring their safety alignment has become critical. Counterintuiti

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

We benchmarked Opus 5 comprehensively on document understanding through ParseBench. It is roughly on par with Opus 4.8 - it does a few point…

DGX agent

We benchmarked Opus 5 comprehensively on document understanding through ParseBench. It is roughly on par with Opus 4.8 - it does a few points worse on dense tables, but does slightly better on parsing

model-releasesjerry-liu--x
24 Jul 2026
Model Releases

A Multiclass Quantum Aligned Centroid Kernel

DGX agent

arXiv:2607.19782v1 Announce Type: cross Abstract: Kernel methods are powerful tools in machine learning but commonly used full-Gram kernels face three key limitations: (1) quadratic scaling with train

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

A Systematic Evaluation of Traditional Privacy Policy Analysis Tools Against LLMs

DGX agent

arXiv:2607.17075v2 Announce Type: replace-cross Abstract: The advent of LLMs has significantly changed the research on privacy policy and data compliance analysis by enabling tasks that previously req

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

A Unified Tokenization Framework for Pain Recognition using Heterogeneous 3D Modalities

DGX agent

arXiv:2607.19716v1 Announce Type: new Abstract: Pain is a complex and pervasive phenomenon affecting a large percentage of the population, and accurate assessment is essential for effective clinical m

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

AgentCgroup: Understanding and Controlling OS Resources of AI Agents

DGX agent

arXiv:2602.09345v3 Announce Type: replace-cross Abstract: AI agents are increasingly deployed in multi-tenant cloud environments, where they execute diverse tool calls within sandboxed containers, eac

model-releasesarxiv-cs-ai
23 Jul 2026
← Previous
1…597598599600601…1395
Next →