AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
All
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,545 results
Model Releases

The Gemma series is an amazing set of highly performant open-weight models. They have proven extremely effective in industrial settings wher…

DGX agent

The Gemma series is an amazing set of highly performant open-weight models. They have proven extremely effective in industrial settings where site-deployed agents need exactly this as a base for domai

model-releasesyann-lecun--x
25 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

The Opus 5 system card itself is a fun PDF to parse. It's 193 pages and stacked with labeled and unlabeled charts 📊 LlamaParse does a surpr…

DGX agent

The Opus 5 system card itself is a fun PDF to parse. It's 193 pages and stacked with labeled and unlabeled charts 📊 LlamaParse does a surprisingly good job on agentic (1.25c per page) and agentic plus

model-releasesjerry-liu--x
25 Jul 2026
Model Releases

Very happy to support this on behalf of Google. We have long benefited from open source, are big contributors to open source and in fact hav…

DGX agent

Very happy to support this on behalf of Google. We have long benefited from open source, are big contributors to open source and in fact have consistently made open weights models with Gemma available

model-releasesclem-delangue--x
25 Jul 2026
Model Releases

A Comparative Evaluation of Embeddings and LLMs in a Greek Book Publisher Setting - The CUP Dataset

DGX agent

arXiv:2607.21274v1 Announce Type: cross Abstract: We present CUP, a Greek book retrieval benchmark consisting of 868 catalog records and 104 expert-annotated queries with graded relevance judgments. W

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

A Knowledge-Injection Framework for Zero-Shot Adaptation of LLMs to Delirium Prediction

DGX agent

arXiv:2607.20453v1 Announce Type: cross Abstract: Large language models show promise for clinical prediction, but zero-shot performance on specialized tasks is limited by incomplete domain knowledge,

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

A new model launch is not a product update ‼ It's one of three things, and you don't know which until you test it. Sometimes it's nothing: t…

DGX agent

A new model launch is not a product update ‼ It's one of three things, and you don't know which until you test it. Sometimes it's nothing: the model improved inside the same distribution, your harness

model-releasesitamar-friedman--x
24 Jul 2026
Model Releases

A reminder this is 20% off via the Nous Portal <3

DGX agent

mr‑r0b0t announced that Claude Opus 5 is available at a 20 % discount through the Nous Portal. The model can be accessed via the Hermes Agent on the Nous Portal, as well as through OpenRouter and Anth

model-releasesnous-research--x
24 Jul 2026
Model Releases

A Sovereign, Open-Source Foundation Model for German and English

DGX agent

arXiv:2607.09424v3 Announce Type: replace-cross Abstract: We present Soofi S 30B-A3B, a sovereign, open-source Mixture-of-Experts (MoE) hybrid Mamba Transformer foundation model for German and English

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Achieving Text-based Person Retrieval with Any Granularity

DGX agent

arXiv:2607.21057v1 Announce Type: new Abstract: Text-based person retrieval faces a critical but under-explored challenge: the inherent uncertainty of query granularity in real-world scenarios. This p

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

ADABORD: a novel AdaBoost approach for ordinal classification

DGX agent

arXiv:2607.21003v1 Announce Type: new Abstract: Ordinal Classification (OC) deals with classification tasks where the classes follow a natural order. Despite the progress in OC, many existing approach

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Adaptive Multi-Horizon Reinforcement Learning

DGX agent

arXiv:2607.20656v1 Announce Type: cross Abstract: Effective decision-making in complex and changing environments requires balancing short-term and long-term consequences. In reinforcement learning (RL

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Agentic coding without the cloud: evaluating open-weight large language models on longitudinal data preparation tasks

DGX agent

arXiv:2607.21482v1 Announce Type: new Abstract: Large language models (LLMs) and agents are now widely used tools in code development, with data typically sent to third-party cloud-based models. Their

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Agentic Designer: Progressive Multi-Agent Collaboration for Structure-Aware Interior Layout Generation

DGX agent

arXiv:2607.20866v1 Announce Type: new Abstract: Generating realistic interior furniture layouts that strictly adhere to architectural constraints (e.g., walls, doors, and windows) remains a fundamenta

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

AI Assistants Overassist

DGX agent

arXiv:2607.21306v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as tutors and thought partners, helping users reason through problems. While guidance from AI assis

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

AISE-Bench: A Full-Cycle Curated Benchmark for Information Seeking on Academic Knowledge Graphs

DGX agent

arXiv:2607.20498v1 Announce Type: new Abstract: Large language models (LLMs) augmented with tools are emerging as autonomous agents capable of using Web engine, APIs, and code to solve complex, long-h

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Am I supposed to interpret that it's better than Fable 5 from the benchmarks? Or it's ~close but cheaper?

DGX agent

Am I supposed to interpret that it's better than Fable 5 from the benchmarks? Or it's ~close but cheaper? Introducing Claude Opus 5. It's a thoughtful and proactive model that comes close to the front

model-releasesjerry-liu--x
24 Jul 2026
Model Releases

An Analytically Trained Variational Surrogate for Quantum Phase Estimation on NISQ Hardware

DGX agent

arXiv:2607.20943v1 Announce Type: cross Abstract: Quantum Phase Estimation (QPE) is a foundational algorithm for molecular ground-state energy estimation, but its deep circuit requirements make direct

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

An interesting thing I'm observing from the blog/system card is that on a good chunk of the reported benchmarks (~20-30% from a skim), Opus …

DGX agent

An interesting thing I'm observing from the blog/system card is that on a good chunk of the reported benchmarks (~20-30% from a skim), Opus 5 max thinking leads to a degradation in performance compare

model-releasesjerry-liu--x
24 Jul 2026
Model Releases

An LLM-Driven Workflow for Automated Process Control Strategy Generation and Tuning from Dynamic Process Models

DGX agent

arXiv:2607.21292v1 Announce Type: new Abstract: We present a structured large-language-model-driven workflow for automated multi-variable control design from dynamic process models. The workflow decom

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Announcing Fugu-Ultra v1.1 🐡 We’ve been thrilled by the reception to the Fugu model family. Thanks to everyone who tried it, shared feedbac…

DGX agent

Announcing Fugu-Ultra v1.1 🐡 We’ve been thrilled by the reception to the Fugu model family. Thanks to everyone who tried it, shared feedback, and trusted Fugu with real work. Today, we’re releasing Fu

model-releasesdavid-ha--x
24 Jul 2026
Model Releases

Anthropic launches Claude Opus 5, which it says comes close to Fable 5 performance at half the price; it is the new default model on Claude Max (Anthropic)

DGX agent

Anthropic: Anthropic launches Claude Opus 5, which it says comes close to Fable 5 performance at half the price; it is the new default model on Claude Max — Claude Opus 5 is available today. It's a th

model-releasestechmeme
24 Jul 2026
Model Releases

Anthropic releases Opus 5 with ‘close’ to Fable 5’s capabilities

DGX agent

Weeks after Anthropic's latest toe-to-toe with the US government, and days after an OpenAI security incident that dominated tech industry discussions, Anthropic on Thursday released its newest model,

model-releasesthe-verge-ai
24 Jul 2026
Model Releases

Anthropic says Opus 5 is the company's 'most aligned model to date'; it is Anthropic's fourth model release in less than two months (Madison Mills/Axios)

DGX agent

Madison Mills / Axios: Anthropic says Opus 5 is the company's “most aligned model to date”; it is Anthropic's fourth model release in less than two months — Anthropic on Thursday is releasing Claude O

model-releasestechmeme
24 Jul 2026
Model Releases

AppWorld-UL: Benchmarking Diverse Agent-User Interactions for Tool-Use

DGX agent

arXiv:2607.20536v1 Announce Type: new Abstract: Tool-use agents that address day-to-day digital tasks such as ordering groceries must not only operate applications, but also interact with the user, e.

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

ArbiGraph: Arbitrarily Scalable Verifiable Task Graphs for Evaluating Context Management

DGX agent

arXiv:2607.20764v1 Announce Type: new Abstract: We introduce ARBIGRAPH, a benchmark generator for evaluating whether tool-assisted language agents can retain, update, compose, and discard task-relevan

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Are Single-Token Sparse Autoencoder Features Causally Necessary? Layer-Depth and SAE-Family Effects

DGX agent

arXiv:2607.20596v1 Announce Type: cross Abstract: Sparse autoencoder (SAE) features are used to interpret and steer large language models, yet whether a feature's causal role is stable across SAE fami

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

AREX: Towards a Recursively Self-Improving Agent for Deep Research

DGX agent

arXiv:2607.21461v1 Announce Type: new Abstract: Deep research requires agents to find answers that jointly satisfy multiple constraints. Discovering such answers is costly, whereas verifying a candida

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

As a joke I prompted Codex 'Build and run BenchBench, a benchmark of now good ai is at creating benchmarks. then figure out what benchbenchb…

DGX agent

As a joke I prompted Codex 'Build and run BenchBench, a benchmark of now good ai is at creating benchmarks. then figure out what benchbenchbench is and run that. and then write benchbenchbench up as a

model-releasesethan-mollick--x
24 Jul 2026
Model Releases

Attention Degradation, Function Token Anchoring, and the Limits of Attention-Based Intervention in Large Language Models

DGX agent

arXiv:2607.20524v1 Announce Type: new Abstract: Mean cross-positional attention degradation is widely reported in transformer interpretability, yet whether it causally limits contextual retrieval rema

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Attribution Markets: A Fisher-Market Formulation for Fractional Credit Assignment Between Planned Tasks and Performed Actions

DGX agent

arXiv:2607.20694v1 Announce Type: new Abstract: Personal and organizational planning systems maintain two records that drift apart: what was planned (a task's effort budget) and what was done (a logge

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

[audio.cpp] Release 0.4: Higgs Audio v3 TTS 4B (10x real time)+ Fish Audio S2 Pro in C++/GGML, full GGUF loading, Q8 speed and VRAM gains

DGX agent

audio.cpp again :) Release 0.4 is out. The headline this time is new high-quality TTS coverage plus GGUF becoming a first-class across the project. What’s new: Added Higgs Audio v3 TTS 4B, Fish Audio

model-releasesr-localllama
24 Jul 2026
Model Releases

Automated Synthesis and Adversarial Validation of Executable Causal Research Pipelines

DGX agent

arXiv:2607.21173v1 Announce Type: new Abstract: While automated research systems promise to accelerate empirical analysis, they are prone to silent failures: instances in which analysis code executes

model-releasesarxiv-cs-lg
24 Jul 2026
Model Releases

Autonomous disproofs of the sum-product conjecture over mathbb R with GPT-5.5 Pro

DGX agent

arXiv:2607.20525v1 Announce Type: new Abstract: OpenAI's recent disproof of the Erdos unit distance conjecture marked a milestone for AI in mathematics. It also inspired another breakthrough: a human

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Autonomous Topology Mutation: Safe Runtime Restructuring for Multi-Agent LLM Systems with Capability, State, and Shadow Invariants

DGX agent

arXiv:2607.20488v1 Announce Type: new Abstract: Multi-agent LLM frameworks typically fix their team topology at boot time. When an individual agent becomes overloaded at runtime, for example by mixing

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

AXIS: A Growable Community-Driven Data Engine for Scalable Robot Manipulation

DGX agent

arXiv:2607.21588v1 Announce Type: new Abstract: Learning effective robot manipulation policies requires diverse, high-quality demonstrations, yet existing data pipelines are often difficult to scale b

model-releasesarxiv-cs-ro
24 Jul 2026
Model Releases

b10103

DGX agent

metal : add f16 type support to leaky relu (#25981) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFra

model-releasesllama-cpp-releases
24 Jul 2026
Model Releases

b10105

DGX agent

args: refactor mlock/mmap/directio into load-mode (#20834) args: overhaul mmap/mlock/dio into single arg Signed-off-by: Aaron Teo aaron.teo1@ibm.com docs: update docs with llama-gen-docs Signed-off-by

model-releasesllama-cpp-releases
24 Jul 2026
Model Releases

b10106

DGX agent

CUDA: fix external compilation of q1_0 MMQ (#25778) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFra

model-releasesllama-cpp-releases
24 Jul 2026
Model Releases

b10107

DGX agent

hexagon: fix Windows crash when op_poll is enabled (#26029) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) i

model-releasesllama-cpp-releases
24 Jul 2026
Model Releases

Benchmarking Large Language Models on Multi-Sensor Physical Hazard Assessment

DGX agent

arXiv:2607.20476v1 Announce Type: new Abstract: We present an empirical benchmark evaluating how five large language models assess multisensor physical hazard data. Testing 60 scenarios across three c

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Benchmarking the Personalization Capabilities of Large Language Models

DGX agent

arXiv:2607.20471v1 Announce Type: new Abstract: Personalization, the act of varying a message to induce action from a specific receiver while keeping sender, channel, and time fixed, has a long tradit

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Benchmarking Unlearning for Vision Transformers

DGX agent

arXiv:2602.20114v2 Announce Type: replace-cross Abstract: Machine unlearning (MU) refers to the post-training capability to remove (the influence of) training examples that are incorrect, biased, or l

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Beyond Liars' Bench: The Impact of Lie Typology, Depth, and Sparsity on Deception Detection in LLMs

DGX agent

arXiv:2607.20479v1 Announce Type: new Abstract: Training probes to detect deceptive outputs from large language models is still an open problem. Recent work has demonstrated that detection probes fail

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Break Through the Compression Bottleneck: From Theory to Practice

DGX agent

arXiv:2607.20434v1 Announce Type: cross Abstract: As the parameter size of language models continues to grow, effective model compression is required to reduce their computational and memory overhead.

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

CachyLLama’s: llama.cpp fork with persistent KV cache that makes long local-agent sessions much less painful

DGX agent

I’m not affiliated with this project, but I’ve been running it recently and I’m surprised it hasn’t received more attention here: https://github.com/fewtarius/CachyLLama CachyLLama is a fork of llama.

model-releasesr-localllama
24 Jul 2026
Model Releases

CAMeR: Keyword-Gated Hybrid Activation for Adaptive Memory Retention in LLM Agents

DGX agent

arXiv:2607.20458v1 Announce Type: cross Abstract: Large language model (LLM) agents operating over extended dialogues accumulate vast amounts of information, yet existing memory systems either retain

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Can LLMs solve mazes?

DGX agent

https://reddit.com/link/1v5rvuq/video/bgmwc754i9fh1/player My goal was to create a benchmark to measure the spatial awareness and memory of models. Eventually, I came up with the simple idea of a maze

model-releasesr-localllama
24 Jul 2026
Model Releases

CANN Bench: Benchmarking Agent Generated Kernels against Real NPU and Algorithmic Limits

DGX agent

arXiv:2607.20518v1 Announce Type: new Abstract: AI agents are now capable of writing, compiling, and iteratively optimizing low-level operator kernels on different hardware platforms. Existing benchma

model-releasesarxiv-cs-ai
24 Jul 2026
← Previous
1…8687888990…470
Next →