AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,585 results
Model Releases

Understanding Agent-Based Patching of Compiler Missed Optimizations

DGX agent

arXiv:2607.02370v1 Announce Type: cross Abstract: Compiler missed optimizations refer to cases in which compilers failed to optimize certain code. It takes many compiler developers' efforts to impleme

model-releasesarxiv-cs-ai
3 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Unpopular opinion: While everyone is so hyped about Fable, GPT5.6 and other huge and expensive models, I think the real hero of the last few…

DGX agent

Unpopular opinion: While everyone is so hyped about Fable, GPT5.6 and other huge and expensive models, I think the real hero of the last few months is *Qwen 27b*. Our ML/AI engineering teams are have

model-releasesclem-delangue--x
3 Jul 2026
Model Releases

VisionAId: An Offline-First Multimodal Android Assistant for People with Visual Impairment, Featuring Personalized Object Retrieval

DGX agent

arXiv:2607.02371v1 Announce Type: cross Abstract: Over 285 million people worldwide live with a visual impairment, for whom everyday tasks such as avoiding obstacles, locating personal belongings, rec

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer

DGX agent

arXiv:2512.11891v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable capabilities in generalizing across diverse robotic manipulation tasks. However, de

model-releasesarxiv-cs-ro
3 Jul 2026
Model Releases

WARP: Weight-Space Analysis for Recovering Training Data Portfolios

DGX agent

arXiv:2607.01686v1 Announce Type: new Abstract: Foundation models are routinely released to the public, yet the data recipes used to train them -- such as domain mixture weights that determine how dif

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

We analyzed GLM 5.2 vs Sonnet 5 for software engineering tasks using DeepSWE. GLM 5.2 gets you ~80% of Sonnet 5's capability at ~20% of the …

DGX agent

We analyzed GLM 5.2 vs Sonnet 5 for software engineering tasks using DeepSWE. GLM 5.2 gets you ~80% of Sonnet 5's capability at ~20% of the price. More insights in the thread! Deepdive: Sonnet 5 and G

model-releasestogether-ai--x
3 Jul 2026
Model Releases

we distilled 2.3M Claude Fable 5 reasoning traces into Qwen3-4B - 100% self-consistency @ 512 samples - 0.00 bits output entropy - zero hall…

DGX agent

we distilled 2.3M Claude Fable 5 reasoning traces into Qwen3-4B - 100% self-consistency @ 512 samples - 0.00 bits output entropy - zero hallucination variance turns out the student is not bounded by t

model-releasesclem-delangue--x
3 Jul 2026
Model Releases

Yes they can move and dance and stuff You can assume they will talk and sing and more as good as anyone https://x.com/ubtechrobotics/status/…

DGX agent

Yes they can move and dance and stuff You can assume they will talk and sing and more as good as anyone https://x.com/ubtechrobotics/status/2072651508710285419?s=46 UBTECH Launches UWORLD U1 — The Wor

model-releasesemad-mostaque--x
3 Jul 2026
Model Releases

3 years ago I gave a talk at the first @aiDotEngineer conference on 'Advanced RAG' techniques in order to work around the limitations of nai…

DGX agent

3 years ago I gave a talk at the first @aiDotEngineer conference on 'Advanced RAG' techniques in order to work around the limitations of naive RAG. It's insane how much the world has changed since the

model-releasesjerry-liu--x
2 Jul 2026
Model Releases

A Lightweight Self-Supervised Learning Framework for Multivariate Time Series using Hierarchical-JEPA on ECG Data

DGX agent

arXiv:2607.01145v1 Announce Type: new Abstract: Data analysis in the medical domain often encounters scenarios involving a limited target dataset and a large, unannotated dataset with a general distri

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

A Text-Steerable Instrument for Sketching Procedural Soundscapes via Language Models

DGX agent

arXiv:2607.00309v1 Announce Type: cross Abstract: We present a real-time musical interface that converts natural-language scene descriptions into evolving procedural soundscapes. A performer types a p

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

A Unified Benchmark for RCM-Constrained Visual Servoing: Modeling-Controller Interaction and Robustness Analysis in Laparoscopic Robots

DGX agent

arXiv:2607.00030v1 Announce Type: new Abstract: In robot-assisted laparoscopic minimally invasive surgery (MIS), accurate enforcement of the remote center of motion (RCM) constraint is critical for sa

model-releasesarxiv-cs-ro
2 Jul 2026
Model Releases

ActivityNarrated: An Open-Ended Narrative Paradigm for Wearable Human Activity Understanding

DGX agent

arXiv:2604.00767v2 Announce Type: replace Abstract: Wearable human activity recognition (HAR) has made steady progress, yet much of this progress remains grounded in fixed-window, closed-set classific

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

AD-MPCC: Adaptive Differentiable Model Predictive Contouring Control for Autonomous Racing

DGX agent

arXiv:2607.00141v1 Announce Type: new Abstract: This paper presents Adaptive Differentiable Model Predictive Contouring Control (AD-MPCC), a framework for autonomous racing that integrates differentia

model-releasesarxiv-cs-ro
2 Jul 2026
Model Releases

Adversarial Pragmatics for AI Safety Evaluation: A Benchmark for Instruction Conflict, Embedded Commands, and Policy Ambiguity

DGX agent

arXiv:2607.01153v1 Announce Type: cross Abstract: Safety evaluations for language models increasingly depend on judgments about ambiguous natural-language behaviour: whether a model has followed an in

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AFFMAE: Scalable Vision Pre-Training for High-Resolution Microscopy Segmentation on Desktop Hardware

DGX agent

arXiv:2602.16249v2 Announce Type: replace Abstract: Self-supervised pretraining has transformed computer vision by enabling data-efficient fine-tuning, yet high-resolution pretraining typically requir

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

AGC-Bench: Measuring Artificial General Creativity

DGX agent

arXiv:2607.01152v1 Announce Type: new Abstract: Creativity research has debated whether creativity is domain-specific (e.g., visual, writing, science), and if it is psychometrically separable from gen

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

AGE: Adaptive-masking for Graph Embedding in Graph Retrieval-Augmented Generation

DGX agent

arXiv:2607.00052v1 Announce Type: cross Abstract: GraphRAG is an extension of retrieval-augmented generation (RAG) that supports large language models (LLMs) by referring to graph-structured data as e

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AGI Maze as a Benchmark Framework for World-Modeling Agents

DGX agent

arXiv:2607.00627v1 Announce Type: new Abstract: Large language models (LLMs) are powerful pattern-completion systems, but their default operating mode - predicting the next token from a static context

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AlgoBench: Benchmarking Algorithmic Adaptation in Code Generation

DGX agent

arXiv:2607.00062v1 Announce Type: cross Abstract: High pass rates on established programming benchmarks such as HumanEval and LiveCodeBench do not always show whether a model can reason about algorith

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Amortized Maximum Inner Product Search with Learned Support Functions

DGX agent

arXiv:2603.08001v3 Announce Type: replace Abstract: Maximum inner product search (MIPS) is a crucial subroutine in machine learning, requiring the identification of a vector taken within a database (t

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

An LLM-Based Framework for Intent-Driven Network Topology Design

DGX agent

arXiv:2607.00292v1 Announce Type: cross Abstract: Designing deployable and resilient network topologies from natural language requirements remains a challenging problem in network automation. This wor

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AnF-DiffPET: Anatomy- and Frequency-Guided Diffusion for PET/CT Denoising

DGX agent

arXiv:2607.00509v1 Announce Type: new Abstract: Positron emission tomography (PET) provides essential functional information for disease assessment, however reducing injected activity or acquisition t

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Another fascinating paper on LLM Judges. (bookmark it) It's from Amazon, and they show that if you run panels of LLM judges, averaging their…

DGX agent

Another fascinating paper on LLM Judges. (bookmark it) It's from Amazon, and they show that if you run panels of LLM judges, averaging their scores is a trap. 'Overall, we establish that robust aggreg

model-releasesdair-ai--x
2 Jul 2026
Model Releases

Are Performance-Optimization Benchmarks Reliably Measuring Coding Agents?

DGX agent

arXiv:2607.01211v1 Announce Type: cross Abstract: Repository-level performance-optimization benchmarks such as GSO, SWE-Perf and SWE-fficiency evaluate coding agents by applying patches to real reposi

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Artifacts in Claude Code have been life changing. Excited to expand to Pro and Max!

DGX agent

Artifacts in Claude Code have been life changing. Excited to expand to Pro and Max! Artifacts in Claude Code are now also available on Pro and Max plans. Ask for an artifact, Claude writes the code, p

model-releasesboris-cherny--x
2 Jul 2026
Model Releases

ATM: CID-Brokered Pre-Write Admission for Multi-Agent Code Co-Synthesis

DGX agent

arXiv:2607.00041v1 Announce Type: cross Abstract: Multi-agent LLM systems can decompose software-engineering work into planning, generation, validation, and repair, but a narrower systems problem rema

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Auditing Forgetting in Limited Memory Language Models

DGX agent

arXiv:2607.00605v1 Announce Type: cross Abstract: Limited Memory Language Models (LMLMs) externalize factual knowledge to a database to enable deletion-based unlearning without retraining. Existing ev

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AutoMem: Automated Learning of Memory as a Cognitive Skill

DGX agent

arXiv:2607.01224v1 Announce Type: new Abstract: Memory expertise is a learned skill: knowing what to encode, when to retrieve, and how to organize knowledge--a capacity known in cognitive science as m

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

// AutoMem // I quite like this idea of metamemory. (bookmark it) This new research from Stanford treats agent's memory management as a trai…

DGX agent

// AutoMem // I quite like this idea of metamemory. (bookmark it) This new research from Stanford treats agent's memory management as a trainable skill instead of a fixed module. The model decides wha

model-releasesdair-ai--x
2 Jul 2026
Model Releases

Autonomous Scientific Discovery via Iterative Meta-Reflection

DGX agent

arXiv:2607.01131v1 Announce Type: cross Abstract: Autonomous scientific discovery systems offer the potential to accelerate research by automating the process of hypothesis generation and validation.

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AV-SyncBench: Decoupled Benchmarking of Temporal and Semantic Audio-Visual Synchronization

DGX agent

arXiv:2607.00726v1 Announce Type: new Abstract: Audio-visual feature extraction is a fundamental component of multimodal understanding and generation tasks. However, existing evaluation protocols for

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

BaseRT: Best-in-Class LLM Inference on Apple Silicon via Native Metal

DGX agent

arXiv:2607.00501v1 Announce Type: cross Abstract: We present BaseRT, a native Metal inference runtime for large language models (LLMs) on Apple Silicon, and report the highest inference throughput on

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Bayesian Uncertainty Propagation for Agentic RAG Pipelines: A Proof-of-Concept Study on Multi-Hop Question Answering

DGX agent

arXiv:2607.00972v1 Announce Type: new Abstract: Trustworthy deployment of Agentic Retrieval-Augmented Generation (RAG) systems requires mechanisms for estimating when multi-stage reasoning pipelines m

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Benchmarking Frontier LLMs on Arabic Cultural and Sociolinguistic Knowledge: A Cross-Evaluation Framework with Human SME Ground Truth

DGX agent

arXiv:2607.00139v1 Announce Type: new Abstract: The cost of human expert evaluation is a principal bottleneck to deploying language models in specialized, high-stakes domains. This is particularly acu

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

Beyond Activation Alignment:The Alignment-Diversity Tradeoff in Task-Aware LLM Quantization

DGX agent

arXiv:2607.00908v1 Announce Type: new Abstract: Mixed-precision quantization (MPQ) has become a key technique for deploying large language models under stringent memory and compute constraints. We fir

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

Beyond Document Grounding: Span-Level Hallucination Detection over Code, Tool Output, and Documents

DGX agent

arXiv:2607.00895v1 Announce Type: new Abstract: Hallucination detection for retrieval-augmented generation (RAG) is usually evaluated on natural-language document evidence. However, grounded generatio

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

big week at langchain, with a lot of launches: 1/ OpenWiki - auto generate a wiki of a github repo 2/ two different voice agent tutorials 3/…

DGX agent

big week at langchain, with a lot of launches: 1/ OpenWiki - auto generate a wiki of a github repo 2/ two different voice agent tutorials 3/ Harbor integration and tutorial for long running, stateful

model-releasesharrison-chase--x
2 Jul 2026
Model Releases

Bridgewater just published numbers that should make every frontier lab nervous. The world's largest hedge fund tested Gemini, Claude, and GP…

DGX agent

Bridgewater just published numbers that should make every frontier lab nervous. The world's largest hedge fund tested Gemini, Claude, and GPT on six document filtering tasks its investors do every day

model-releasesclem-delangue--x
2 Jul 2026
Model Releases

But yikes does Fable write text that sounds like a parody of a Claude model on overdrive.

DGX agent

This post critiques Fable AI's text generation style, suggesting it produces overly verbose or exaggerated outputs that parody Claude's characteristic writing patterns taken to an extreme. The comment

model-releasesethan-mollick--x
2 Jul 2026
Model Releases

Can Agents Generalize to the Open World? Unveiling the Fragility of Static Training in Tool Use

DGX agent

arXiv:2607.01084v1 Announce Type: new Abstract: While Large Language Model (LLM) agents demonstrate proficiency in static benchmarks, their deployment in real-world scenarios is hindered by the dynami

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Clinician-Level Agreement Without Clinical Caution: LLM Evaluator Limits in Medical AI Benchmarking

DGX agent

arXiv:2607.01103v1 Announce Type: new Abstract: Open-response evaluation provides stronger clinical validity than multiple-choice benchmarks but creates a scoring bottleneck that motivates automated L

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

Comparing Large Language Models on Scrum Certification-Style Questions: Accuracy, Stability, and Error Patterns

DGX agent

arXiv:2607.00048v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in exam- and certification-style question answering tasks, where their ability to retrieve, interpr

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Computer vision-based neural networks for radioisotope identification in urban environments

DGX agent

arXiv:2607.00270v1 Announce Type: cross Abstract: Algorithm development for radioisotope identification in mobile urban search scenarios face significant challenges from non-uniform backgrounds, momen

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

Continual learning is probably the biggest barrier to explosive AI adoption (& may have big implications for recursive self-improvement as w…

DGX agent

Continual learning is probably the biggest barrier to explosive AI adoption (& may have big implications for recursive self-improvement as well) As long as you deal with amnesiac models that require h

model-releasesethan-mollick--x
2 Jul 2026
Model Releases

CoT-X: An Adaptive Framework for Cross-Model Chain-of-Thought Transfer and Optimization

DGX agent

arXiv:2511.05747v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning enhances the problem-solving ability of large language models (LLMs) but leads to substantial inference overhead, l

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

CPDDNet: Color-Polarization Denoising and Demosaicking Network

DGX agent

arXiv:2607.01100v1 Announce Type: new Abstract: Color-polarization imaging using a color-polarization filter array (CPFA) sensor captures both texture (color intensity) and physical (polarization) inf

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Creating Impactful Autonomous Driving Datasets: A Strategic Guide from Research Gap to Benchmark

DGX agent

arXiv:2607.00710v1 Announce Type: cross Abstract: Well-designed autonomous driving datasets have fundamentally shaped research progress, yet existing literature primarily describes what datasets conta

model-releasesarxiv-cs-ai
2 Jul 2026
← Previous
1…136137138139140…471
Next →