AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
All
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,553 results
Model Releases

Benchmarking the Personalization Capabilities of Large Language Models

DGX agent

arXiv:2607.20471v1 Announce Type: new Abstract: Personalization, the act of varying a message to induce action from a specific receiver while keeping sender, channel, and time fixed, has a long tradit

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Benchmarking Unlearning for Vision Transformers

DGX agent

arXiv:2602.20114v2 Announce Type: replace-cross Abstract: Machine unlearning (MU) refers to the post-training capability to remove (the influence of) training examples that are incorrect, biased, or l

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Beyond Liars' Bench: The Impact of Lie Typology, Depth, and Sparsity on Deception Detection in LLMs

DGX agent

arXiv:2607.20479v1 Announce Type: new Abstract: Training probes to detect deceptive outputs from large language models is still an open problem. Recent work has demonstrated that detection probes fail

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Break Through the Compression Bottleneck: From Theory to Practice

DGX agent

arXiv:2607.20434v1 Announce Type: cross Abstract: As the parameter size of language models continues to grow, effective model compression is required to reduce their computational and memory overhead.

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

CachyLLama’s: llama.cpp fork with persistent KV cache that makes long local-agent sessions much less painful

DGX agent

I’m not affiliated with this project, but I’ve been running it recently and I’m surprised it hasn’t received more attention here: https://github.com/fewtarius/CachyLLama CachyLLama is a fork of llama.

model-releasesr-localllama
24 Jul 2026
Model Releases

CAMeR: Keyword-Gated Hybrid Activation for Adaptive Memory Retention in LLM Agents

DGX agent

arXiv:2607.20458v1 Announce Type: cross Abstract: Large language model (LLM) agents operating over extended dialogues accumulate vast amounts of information, yet existing memory systems either retain

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Can LLMs solve mazes?

DGX agent

https://reddit.com/link/1v5rvuq/video/bgmwc754i9fh1/player My goal was to create a benchmark to measure the spatial awareness and memory of models. Eventually, I came up with the simple idea of a maze

model-releasesr-localllama
24 Jul 2026
Model Releases

CANN Bench: Benchmarking Agent Generated Kernels against Real NPU and Algorithmic Limits

DGX agent

arXiv:2607.20518v1 Announce Type: new Abstract: AI agents are now capable of writing, compiling, and iteratively optimizing low-level operator kernels on different hardware platforms. Existing benchma

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Capital Markets LLM Reliability Score (CM-LRS): From Plausible to Bankable

DGX agent

arXiv:2607.21340v1 Announce Type: new Abstract: In capital-markets workflows the question is rarely whether a large language model can produce a fluent draft, but whether the draft is bankable: defens

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

Cardinality-Decomposed Loss: Matching Training Objectives to Relation Structure in Heterogeneous Recommendation Graphs

DGX agent

arXiv:2607.20737v1 Announce Type: new Abstract: Graph Neural Networks trained on heterogenous bipartite graphs form a common basis in recommendation systems. These graphs often express relations that

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Case study: solving P-99 with LPTP and an LLM

DGX agent

arXiv:2607.21196v1 Announce Type: cross Abstract: Ninety-Nine Prolog Problems (P-99) is a famous set of Prolog exercises. We solved the first thirty three just by prompting an LLM (Large Language Mode

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Chronofy: A Temporal-Logical Decay Architecture for Information Validity in Time-Aware Retrieval-Augmented Generation

DGX agent

arXiv:2607.20560v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems retrieve and integrate external knowledge to ground large language model (LLM) outputs. However, current RA

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Classifier-pruned Bayesian optimization for particle accelerator tuning: Exploring temporally structured manifold of 6D beam phase space

DGX agent

arXiv:2412.01748v2 Announce Type: replace Abstract: Complex dynamical systems, such as particle accelerators, often require intricate and time-consuming tuning procedures to achieve optimal performanc

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

ConfidenceBench: Evaluating Confidence Calibration in Large Language Models

DGX agent

arXiv:2607.20526v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in settings where fluent but incorrect answers can be costly. In these settings, accuracy alone i

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

CRAG-MM-Diagnostics: Enabling Stage-Wise Analysis of Knowledge-Intensive VQA

DGX agent

arXiv:2607.21155v1 Announce Type: cross Abstract: Knowledge-Intensive Visual Question Answering (KI-VQA) benchmarks evaluate Vision-Language Models (VLMs) as multimodal knowledge assistants by requiri

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

CT-Merging: Consensus Directions and Task-Level Scaling for LoRA Adapter Merging

DGX agent

arXiv:2607.20561v1 Announce Type: cross Abstract: LoRA adapters provide an efficient way to specialize a pretrained model for many downstream tasks, but deploying one adapter per task requires adapter

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

CultureTalk-ID: A Multi-Task Dialogue Benchmark for Cultural Commonsense in Indonesian Local Languages

DGX agent

arXiv:2607.21016v1 Announce Type: new Abstract: Culture is lived through conversation, yet existing Indonesian cultural commonsense benchmarks evaluate LLMs on short and isolated prompts, stripping aw

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

Cycle-Consistent and Uncertainty-Aware Neural Surrogates for Tokamak Edge Plasmas

DGX agent

arXiv:2607.21407v1 Announce Type: cross Abstract: The boundary and divertor plasma govern how a tokamak exhausts power and particles, setting heat fluxes, target conditions, and the onset of detachmen

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

DataPrep-Bench: Benchmarking LLMs as Training Data Preparators

DGX agent

arXiv:2607.20465v1 Announce Type: cross Abstract: The quality of training data fundamentally determines the capabilities of large language models (LLMs), yet no unified benchmark exists to measure how

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

DatedGPT: Preventing Lookahead Bias in Large Language Models with Time-Aware Pretraining

DGX agent

arXiv:2603.11838v2 Announce Type: replace Abstract: Large language models pretrained on internet-scale data risk lookahead bias in forecasting tasks, as they may have already seen the true outcome dur

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

Deblurring in the Wild: A Real-World Image Deblurring Dataset from Smartphone High-Speed Videos

DGX agent

arXiv:2506.19445v4 Announce Type: cross Abstract: We introduce the largest real-world image deblurring dataset constructed from smartphone slow-motion videos. Using 240 frames captured over one second

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Demonstrating GenDB: Instance-Optimized and Customized Query Processing Code Generation via LLM Agents

DGX agent

arXiv:2607.20630v1 Announce Type: cross Abstract: Traditional query processing engines require continuous development and extensions to support new techniques and user requirements, and in some cases,

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Detecting Neural Network Failures through Spectral Analysis of Internal Activations

DGX agent

arXiv:2607.20590v1 Announce Type: cross Abstract: Neural network misclassifications exhibit characteristic spectral instability in internal activations that is invisible at the output layer. This phen

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

DFAH-Bench: Benchmarking Observable Agent Instability in Financial Decision-Making

DGX agent

arXiv:2607.20491v1 Announce Type: new Abstract: Standard evaluation benchmarks measure what a tool-using agent decides, not whether it arrives at that decision through the same process each time. We i

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

DINO-VPT: Hierarchical Visual Prompt Tuning for Joint Physical-Digital Face Anti-Spoofing

DGX agent

arXiv:2607.20900v1 Announce Type: new Abstract: With the increasing diversity of spoofing attacks, there is a growing demand for unified Face Anti-Spoofing (FAS) models capable of detecting both physi

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

Do Active SAE Feature Planes Carry More Holonomy? A Preregistered Reversal in Gemma

DGX agent

arXiv:2607.20522v1 Announce Type: new Abstract: This paper tests whether holonomy concentrates on active sparse-autoencoder (SAE) feature planes in Gemma 2 2B, a concrete operationalization of the bro

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Do Pathology Vision-Language Models Truly See Pathology?

DGX agent

arXiv:2607.21065v1 Announce Type: new Abstract: Pathology vision-language models (VLMs) have recently progressed rapidly and are commonly evaluated by answer accuracy on pathology VQA benchmarks. Howe

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

Domyn-Small: A European 10B Reasoning Language Model

DGX agent

arXiv:2607.20448v1 Announce Type: new Abstract: We introduce Domyn-Small, a 10-billion-parameter open-weight reasoning language model released under the MIT license. Domyn-Small is the product of an i

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

DONDO: Open w2v-BERT Speech-Recognition Base Models for African Languages

DGX agent

arXiv:2607.21540v1 Announce Type: new Abstract: We present DONDO, a family of open, permissively licensed automatic speech recognition (ASR) base models for African languages, built on the w2v-BERT 2.

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

Drive As You Like: Multi-Head Diffusion with Reinforcement Learning for Personalized Driving

DGX agent

arXiv:2508.16947v2 Announce Type: replace-cross Abstract: Despite significant progress, imitation learning-based autonomous driving planners remain largely restricted to reproducing high-frequency bia

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Dropping the Anchor: Statistical Context Summarization for Distributed Systems via Pulsar Attention

DGX agent

arXiv:2607.20457v1 Announce Type: cross Abstract: Inference with large language models (LLMs) on long sequences is computationally expensive due to the quadratic complexity of self-attention. Distribu

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

DynamicMCPBench: A Trace-Grounded, Effect-Scored Benchmark for LLM Agents over Live MCP Servers

DGX agent

arXiv:2607.20531v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly deployed over Model Context Protocol (MCP) servers, yet the benchmarks used to evaluate them score th

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Efficient and Interpretable Body-Based Emotion Recognition with Lightweight Temporal Convolutional Networks

DGX agent

arXiv:2607.20820v1 Announce Type: new Abstract: Body-based emotion recognition is important for real-time affective systems, but graph-based skeleton models can be computationally expensive. This pape

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Engine-Native Editable 3D World Reconstruction with Objects and Lighting

DGX agent

arXiv:2607.20889v1 Announce Type: new Abstract: Editable 3D scene creation requires object instances and lights that can be inspected, moved, and imported into standard engines, yet existing single-im

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

Evaluating the Effectiveness of Persona Simulation in Opinion Prediction with GPT-4.1

DGX agent

arXiv:2607.20589v1 Announce Type: new Abstract: Persona simulation involves utilizing large language models (LLMs) to anticipate human choices or interactions based on specific characteristic informat

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

Expectation Alignment of Language Models for Real-World User Expectations

DGX agent

arXiv:2607.20485v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated remarkable performance on standard benchmarks, yet it remains largely unexplored whether they truly meet

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Explainable Deepfake Detection Challenge

DGX agent

arXiv:2607.21007v1 Announce Type: new Abstract: Deepfake detection is moving beyond binary classification decisions toward systems that can also explain the visual evidence supporting those decisions.

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

Extened garlic to run Qwen3.5 35B A3B float8 at 55 tok/s on RTX 5060 Ti

DGX agent

In a previous post (https://www.reddit.com/r/LocalLLaMA/comments/1utefpr/running_qwen3_30b_a3b_at_50_toks_on_rtx_5060_ti/) there seemed to be great demand for bringing in Qwen3.5 35B. Some Gated Delta

model-releasesr-localllama
24 Jul 2026
Model Releases

Factorized Spatio-Temporal Convolutions for Human Pose Estimation from Planar Lidar

DGX agent

arXiv:2607.21309v1 Announce Type: new Abstract: Localizing nearby humans and estimating their facing direction are key capabilities for safe navigation and socially aware human-robot interaction. Many

model-releasesarxiv-cs-ro
24 Jul 2026
Model Releases

Faster IndexTTS-2: Accelerating and Streaming Autoregressive Zero-Shot Text-to-Speech Synthesis on GPUs

DGX agent

arXiv:2607.21042v1 Announce Type: new Abstract: Autoregressive text-to-speech models achieve strong naturalness but suffer from slow inference due to sequential token generation, limiting their deploy

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Fisher Widths: Local Learning Geometry and Anisotropic Recovery

DGX agent

arXiv:2607.20578v1 Announce Type: new Abstract: We study Gaussian-width complexity on statistical manifolds through a pair of functionals: the primal Fisher width w_G(T) = w(G^{1/2}T), induced by the

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Flash EQ-Linear: Accelerating Equivariant Linear Layers via Group-wise Discrete Fourier Transform

DGX agent

arXiv:2607.21271v1 Announce Type: new Abstract: Equivariant networks embed geometric symmetries as structural priors through weight sharing, achieving remarkable parameter efficiency across vision tas

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

FlashPDE: A Drop-In Fused Triton Operator Library for Neural PDE Solvers

DGX agent

arXiv:2607.18020v2 Announce Type: replace Abstract: Physics-Informed Neural Networks (PINNs) solve PDEs by incorporating physical constraints into neural-network training, but large-scale problems are

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

From a Word-Level Dictionary to Sentence-Level Semantics: Multilingual Grievance Labelling with Contextual Models

DGX agent

arXiv:2607.20946v1 Announce Type: new Abstract: Grievance is one of the warning signs analysts look for when assessing threats of violence. It is increasingly measured at scale from online text, most

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

From Attention to Frequency: Integration of Vision Transformer and FFT-ReLU for Enhanced Image Deblurring

DGX agent

arXiv:2511.10806v1 Announce Type: cross Abstract: Image deblurring is vital in computer vision, aiming to recover sharp images from blurry ones caused by motion or camera shake. While deep learning ap

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

From Evaluation to Optimisation: Hierarchy-Aware Training Signals for CWE Prediction in Python

DGX agent

arXiv:2607.21069v1 Announce Type: new Abstract: The original ALPHA benchmark introduced a taxonomy-aware penalty for evaluating CWE-level vulnerability prediction in Python and proposed that the penal

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Frontier Financial Judgement: Can agents tell what might move a stock?

DGX agent

arXiv:2607.20645v1 Announce Type: cross Abstract: We introduce Frontier Financial Judgement, a challenging new benchmark developed in collaboration with professional equity analysts to assess agents'

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Future Rendering neq Future Surface: A Benchmark and Dataset for Dynamic Surface Reconstruction Beyond the Observed Window

DGX agent

arXiv:2607.21471v1 Announce Type: new Abstract: Dynamic-scene reconstruction is almost always evaluated inside the observed time window, yet deployment settings such as AR overlays, robot interaction,

model-releasesarxiv-cs-cv
24 Jul 2026
← Previous
1…8889909192…470
Next →