AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
10,030 results
25 Jul 2026

Llama.cpp now has full MCP support!

Model ReleasesDGX agent

After a long and grueling effort spearheaded by ngxson, llama.cpp now fully supports MCP for all protocols. Over-the-web HTTP servers were already supported in the client (since they don't require any

Old Coder Needs help with New AI Development and wants to get up to speed to understand it all.

Local AiDGX agent

Hi Guys, I'm an old coder and DBA that has been in the field for almost 40 years. More and more the jobs I was doing for work are being taken over by AI and the need for my type of work is diminishing

24 Jul 2026

A Situational Speech Synthesizer for Yoruba: System Design, Phonological Rule Architecture, and Orthographic Extensions for Contour

Research
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2607.18317v2 Announce Type: replace-cross Abstract: We present TTSYoruba, a rule-based concatenative diphone speech synthesizer for Yoruba, deployed at online as part of the YorubaName.com open

CachyLLama’s: llama.cpp fork with persistent KV cache that makes long local-agent sessions much less painful

Model ReleasesDGX agent

I’m not affiliated with this project, but I’ve been running it recently and I’m surprised it hasn’t received more attention here: https://github.com/fewtarius/CachyLLama CachyLLama is a fork of llama.

Chess_db: A framework for working with large chess game datasets

ResearchDGX agent

arXiv:2607.21195v1 Announce Type: cross Abstract: Chess is a two player strategic game that is embedded in classical AI culture as it was once the frontier for intelligent behaviour. There was the sil

ClickGuard: Detecting and Spoiling Clickbait News with Informativeness Measures and Large Language Models

ResearchDGX agent

arXiv:2607.20463v1 Announce Type: new Abstract: This paper presents an AI-driven browser extension that identifies clickbait to help users avoid misleading Internet articles. Moving beyond traditional

Distinguishing Artificial from Authentic: Evaluating LLMs for Detecting LLM-Generated Content

ResearchDGX agent

arXiv:2607.20446v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly used by students to generate natural language responses and program code, there is growing interest in

Domyn-Small: A European 10B Reasoning Language Model

Model ReleasesDGX agent

arXiv:2607.20448v1 Announce Type: new Abstract: We introduce Domyn-Small, a 10-billion-parameter open-weight reasoning language model released under the MIT license. Domyn-Small is the product of an i

EZASP - Facilitating the Usage of ASP

TutorialsDGX agent

arXiv:2603.26863v2 Announce Type: replace-cross Abstract: Answer Set Programming (ASP) is a declarative programming language used for modeling and solving complex combinatorial problems. It has been s

Foundation-model-guided radiogenomic discovery linking cancer genomes to cancer scans

ResearchDGX agent

arXiv:2607.20583v1 Announce Type: cross Abstract: The function of many genes is still unknown, and conventional driver-discovery methods, which rely on how frequently a gene is mutated, cannot assess

Human-Inspired Framework for Robotic Craniotomy: Integrating Multimodal Fusion and Adaptive Trajectory Adjustment

SafetyDGX agent

arXiv:2607.21058v1 Announce Type: new Abstract: Manual craniotomy is a high-risk, skill-dependent procedure associated with surgeon fatigue and potential dural injury. While robotic approaches have im

ICAE-Bench: Evaluating Coding Agents as Interactive Project Builders

Model ReleasesDGX agent

arXiv:2607.21217v1 Announce Type: new Abstract: The recent emergence of vibe-coding workflows is changing what coding agents are expected to do. Instead of merely completing code under fully specified

Interpretable Embeddings with Sparse Autoencoders: A Data Analysis Toolkit

ResearchDGX agent

arXiv:2512.10092v2 Announce Type: replace Abstract: Analyzing large-scale text corpora is a core challenge in machine learning, crucial for tasks like identifying undesirable model behaviors or biases

Learning to Detect UI Principle Violations via Reinforcement Learning

ApplicationsDGX agent

arXiv:2607.20690v1 Announce Type: new Abstract: Small language models and coding agents increasingly generate web front-end code, yet their outputs are typically evaluated primarily for functional cor

LegalCiteTrust: Benchmarking Citation Trustworthiness in Chinese Long-Form Legal Research Reports

Model ReleasesDGX agent

arXiv:2607.20872v1 Announce Type: new Abstract: Long-form legal research reports increasingly rely on LLMs and agentic research systems, but their reliability depends not only on answering the task, b

Meta makes Muse Spark 1.1 available to consumers, debuts new Facebook features

IndustryDGX agent

Meta Platforms Inc. today made its latest large language model available to consumers through its Meta AI chatbot. The update is rolling out alongside several enhancements to Facebook. Meta’s flagship

Multivariate Planar Curves: A Statistical Framework for Shape Analysis in Images

SafetyDGX agent

arXiv:2508.11780v3 Announce Type: replace-cross Abstract: Recent developments in computer vision have made segmented images widely available across many domains, such as medicine, where segmented radi

Open Knowledge format v0.2 tackles agentic trust

Model ReleasesDGX agent

When we introduced the Open Knowledge Format (OKF) in June 2026, we asserted that the context that agents need (table schemas, metric definitions, runbooks) should live in a format, not in a proprieta

People are using Minecraft farms as AI agent benchmarks

Model ReleasesDGX agent

Someone modelled sugarcane farming as an integer program. See, sugarcane only grows next to water. Water costs one tile and can feed at most four cane tiles. The layout therefore becomes a coverage pr

Position Bias is Hidden Behind Ceiling Effects: A Permutation Diagnostic for LLM Benchmarks

Model ReleasesDGX agent

arXiv:2607.20864v1 Announce Type: cross Abstract: Position bias in multiple-choice LLM evaluation is widely cited as a confound in capability comparisons, but published measurements rely on single ans

Position: Natural Language Should Not Fully Replace Formal Languages

ApplicationsDGX agent

arXiv:2607.20432v1 Announce Type: new Abstract: Recent advances in large language models and their widespread adoption have prompted claims that natural language could entirely replace formal language

Post-Hoc Reasoning in Chain of Thought: Decoding and Steering Pre-Committed Answers

ResearchDGX agent

arXiv:2603.01437v2 Announce Type: replace Abstract: As chain of thought (CoT) has become central to scaling reasoning capabilities in large language models (LLMs), it has also emerged as a promising t

Progressive Cramming: Reliable Token Compression and What It Reveals

ResearchDGX agent

arXiv:2607.21231v1 Announce Type: new Abstract: Token cramming compresses sequences into learned embeddings with near-perfect reconstruction, but fixed token budgets and 99% accuracy thresholds leave

Regulating autonomous and agentic AI

SafetyDGX agent

arXiv:2607.21345v1 Announce Type: new Abstract: Regulating activities where regulatees use autonomous and agentic AI is challenging. Regulatory assumptions about regulatee knowledge and control no lon

RUMBA: Russian User Memory Benchmark

Model ReleasesDGX agent

arXiv:2607.21447v1 Announce Type: cross Abstract: The ability to handle long-term memory in LLMs is becoming increasingly critical, yet existing benchmarks remain English-centric and rely on aggregate

SAP shrugs off fears of AI eating software with big earnings beat

ApplicationsDGX agent

Enterprise software giant SAP SE’s stock inched up in late trading today after it shook off concerns that its business might become a victim of the artificial intelligence boom by posting solid result

The Hidden Footprint: Making Storage a First-Class Metric for LLM Agent Evaluation

Model ReleasesDGX agent

arXiv:2607.11149v3 Announce Type: replace Abstract: LLM agent benchmarks measure task completion, reliability, and inference cost, but not the persistent data an agent run leaves on disk, including lo

Toward Continuous Assurance for the Democratization of AI Agent Creation in Industry

Local AiDGX agent

arXiv:2607.21495v1 Announce Type: new Abstract: AI agents are increasingly created inside organizations by non-engineering users through low-code, no-code, and conversational development environments.

Transition-Related Potentials as Markers of Narrative Comprehension in Continuous EEG

ResearchDGX agent

arXiv:2607.20720v1 Announce Type: cross Abstract: Harnessing the potential of electroencephalography (EEG) for brain research is fundamentally limited by intrinsic noise and the diffuse projection of

Unsupervised Metal Artifact Reduction in Dental CBCT using Fine-tuned Cycle-Consistent Adversarial Networks

ResearchDGX agent

arXiv:2607.20977v1 Announce Type: new Abstract: Metal artifacts generated by dental implants significantly degrade cone-beam computed tomography (CBCT) volumes, obscuring critical anatomical structure

When RLVR Shrinks the Reasoning Boundary: Diagnosing Pass@k Inversion

Model ReleasesDGX agent

arXiv:2607.20543v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve one-sample accuracy while making a model worse under repeated sampling. We study thi

Writhe-Based Polymer Link Classification Using Machine Learning

ResearchDGX agent

arXiv:2607.20657v1 Announce Type: cross Abstract: Unique and rapid classification of knots and links is an open mathematical problem that is relevant to a range of (bio)physical systems, including pol

Zagreus-0.4B-por a small open source language model for Portuguese

Model ReleasesDGX agent

mii-llm, an open source AI lab, released Zagreus-0.4B-por, a compact bilingual Portuguese–English language model pretrained entirely from scratch. The model has approximately 400 million parameters an

23 Jul 2026

A Bayesian Framework for Built-in Input Dimension Reduction for Gaussian Process Modeling

ResearchDGX agent

arXiv:2607.19498v1 Announce Type: cross Abstract: Gaussian process (GP) modeling is widely used in computational science and engineering. However, fitting a GP to high-dimensional inputs remains chall

A local-first harness for multi-agent workflows

Local AiDGX agent

Hey all, I’ve been working on this in my spare time and finally feel ready to share it outside my own circles. Arbiter is a single binary for running agents locally. I originally built it because I wa

Agent-Centric Animal Pose Forecasting

AgentsDGX agent

arXiv:2607.19548v1 Announce Type: new Abstract: Understanding animal behavior at an algorithmic level -- what animals attend to, how they form internal models and plans, and how this maps to action --

Agentic Real2Sim: Physics-based World Modeling with Vision-Language Agents

SafetyDGX agent

arXiv:2607.19190v2 Announce Type: replace-cross Abstract: Real-to-sim conversion for robotic interaction with objects remains labor-intensive because it requires more than visual reconstruction: a str

Asymptotically Optimal Regret for Reinforcement Learning without Horizon Dependence

ResearchDGX agent

arXiv:2607.19854v1 Announce Type: new Abstract: We study horizon-free regret minimization for finite-horizon time-homogeneous tabular Markov decision processes with S states, A actions, horizon H, and

AutoVSR: Automatic Visual-to-Symbolic Reasoning for Symbolic Expression Generation from Circuit Schematic

AgentsDGX agent

arXiv:2607.11338v2 Announce Type: replace Abstract: Symbolic expressions can effectively characterize and predict circuit behavior, but deriving them directly from circuit schematics is challenging. T

Building Fast, Evaluating Slow: Pipeline Choices Dominate Autointerpretability Score Variance

ResearchDGX agent

arXiv:2607.19386v1 Announce Type: cross Abstract: Cross-paper comparison of sparse autoencoder (SAE) interpretability often relies on autointerpretability scores. In this evaluation pipeline, a langua

Challenges of Explainability in Continual Learning for Time Series Forecasting

ApplicationsDGX agent

arXiv:2607.19382v1 Announce Type: cross Abstract: Deep learning models have shown strong potential for time series forecasting, yet their deployment in real-world environmental monitoring remains chal

Closing the Lab-to-Store Gap: A Data-Efficient Post-Training and Experience-Driven Learning VLA Framework for Retail Humanoids

Model ReleasesDGX agent

arXiv:2607.20345v1 Announce Type: cross Abstract: Closing the gap between benchmark performance and reliable real-world operation remains a central challenge for Vision-Language-Action (VLA) humanoid

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation

Model ReleasesDGX agent

arXiv:2607.15434v3 Announce Type: replace-cross Abstract: Multi-agent systems routinely place one AI agent in authority over another. When a subordinate refuses a task, the manager chooses the outcome

CreatiPoster: Towards Editable and Controllable Multi-Layer Graphic Design Generation

Model ReleasesDGX agent

arXiv:2506.10890v2 Announce Type: replace Abstract: Graphic design plays a crucial role in both commercial and personal contexts, yet creating high-quality, editable, and aesthetically pleasing graphi

D2VBench: Benchmarking Large Language Models with Value Dilemmas in Daily Scenarios

Model ReleasesDGX agent

arXiv:2607.19834v1 Announce Type: new Abstract: With the wide application of large language models (LLMs) in real-world scenarios, the value implication of their outputs is crucial. However, existing

EvoDRC: A Self-Evolving Agentic Framework for Automated DRC Violation Repair

Model ReleasesDGX agent

arXiv:2607.20019v1 Announce Type: new Abstract: Design rule check (DRC) closure remains a major bottleneck in advanced-node physical design. Although detailed routers are rule-aware, residual design r

FORCE-Bench: A Benchmark, Dataset, and Evaluation Harness for Agentic AI in Enterprise Finance

Model ReleasesDGX agent

arXiv:2607.19409v1 Announce Type: new Abstract: Recent advances in large language models have accelerated deployment of agentic systems in operational finance. Existing benchmarks emphasize measuring

GraphContainer: A Unified Platform for Comparing and Debugging Graph RAG Methods

TutorialsDGX agent

arXiv:2607.19362v1 Announce Type: new Abstract: Graph RAG mitigates hallucinations and stale knowledge in LLMs, particularly for multi-hop question answering. However, existing approaches remain highl

I wrote the latest of my occasional guides to which AI to use right now for non-experts who want to get stuff done. The agentic systems avai…

AgentsDGX agent

I wrote the latest of my occasional guides to which AI to use right now for non-experts who want to get stuff done. The agentic systems available to everyone are getting extremely powerful (even as th

Kwaipilot/KAT-Coder-V2.5-Dev · Hugging Face

Model ReleasesDGX agent

from kwaipilot: Following the release of KAT-Coder-V2.5 in July, we are pleased to release the open-weight version KAT-Coder-V2.5-Dev, an MOE model with a total parameter count of 35B and 3B activated

Ling-3.0-flash, the new MoE model from @AntLingAGI, is now free in Nous Portal for the next week! At 124B parameters and 5.1B active, it's q…

AgentsDGX agent

Ling-3.0-flash, the new MoE model from @AntLingAGI, is now free in Nous Portal for the next week! At 124B parameters and 5.1B active, it's quick to run and built for agent workloads: coding, search, r

Minimize idle accelerators: Native RL job interleaving with co-operative time-slicing in llm-d

SafetyDGX agent

The math behind reinforcement learning (RL) post-training for large language models (LLMs) is notoriously unforgiving. As frontier AI labs push the boundaries of reasoning and coding models using RL p

Multimodal Large Language Models for Remote Sensing Image Understanding: Domain-Specific or General-Purpose?

ResearchDGX agent

arXiv:2607.20284v1 Announce Type: new Abstract: The rapid development of multimodal large language models (MLLMs) has introduced a flexible paradigm for remote sensing image scene understanding (RSISU

NexForge: Scaling Agent Capabilities through Requirement-Driven Task Synthesis for LLMs

Model ReleasesDGX agent

arXiv:2607.14186v4 Announce Type: replace-cross Abstract: Scaling executable agent training data for LLM post-training is bottlenecked by substrate-bound methods that tie task generation to predefined

// Programmatic Memory Enables Long-Horizon Reasoning // Keep the entire interaction log and search it. It works great and beats bespoke mem…

AgentsDGX agent

// Programmatic Memory Enables Long-Horizon Reasoning // Keep the entire interaction log and search it. It works great and beats bespoke memory harnesses on long-horizon tasks. New research introduces

PSA on Laguna S-2.1 - Use the updated chat template and GGUF

Model ReleasesDGX agent

Link to their official GGUF repo: https://huggingface.co/poolside/Laguna-S-2.1-GGUF/tree/main All the GGUFs received this fix 5ish hours ago - correct yarn_attn_factor to 1.0 (llama.cpp derives mscale

Rethinking Uncertainty Evaluation in Large Language Models

ResearchDGX agent

arXiv:2607.19367v1 Announce Type: new Abstract: Calibration is the primary criterion for evaluating LLM confidence, but it is insufficient: it admits trivially incoherent estimators, depends on the ev

Scaling Time Series Classification via XAI-Driven Data Reduction

HardwareDGX agent

arXiv:2607.15774v2 Announce Type: replace-cross Abstract: Explainable AI (XAI) for time series has seen significant algorithmic growth, but its utility in providing measurable performance gains for do

The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Convergent Category Geometry in Small Language Models

Model ReleasesDGX agent

arXiv:2607.16741v2 Announce Type: replace Abstract: Burger et al. (2024) demonstrated that truth representations in large language models are universal across statement polarity but reside within a mu

The Blueprint: How Voicify makes AI-enabled ordering a delight for customers

Model ReleasesDGX agent

Welcome to The Blueprint, a new feature where we highlight how Google Cloud customers are tackling unique and common challenges across industries using the latest AI and cloud technologies. We hope to

← Previous
1…111112113114115…168
Next →