AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,577 results
Model Releases

3 years ago I gave a talk at the first @aiDotEngineer conference on 'Advanced RAG' techniques in order to work around the limitations of nai…

DGX agent

3 years ago I gave a talk at the first @aiDotEngineer conference on 'Advanced RAG' techniques in order to work around the limitations of naive RAG. It's insane how much the world has changed since the

model-releasesjerry-liu--x
2 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

A Lightweight Self-Supervised Learning Framework for Multivariate Time Series using Hierarchical-JEPA on ECG Data

DGX agent

arXiv:2607.01145v1 Announce Type: new Abstract: Data analysis in the medical domain often encounters scenarios involving a limited target dataset and a large, unannotated dataset with a general distri

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

A Text-Steerable Instrument for Sketching Procedural Soundscapes via Language Models

DGX agent

arXiv:2607.00309v1 Announce Type: cross Abstract: We present a real-time musical interface that converts natural-language scene descriptions into evolving procedural soundscapes. A performer types a p

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

A Unified Benchmark for RCM-Constrained Visual Servoing: Modeling-Controller Interaction and Robustness Analysis in Laparoscopic Robots

DGX agent

arXiv:2607.00030v1 Announce Type: new Abstract: In robot-assisted laparoscopic minimally invasive surgery (MIS), accurate enforcement of the remote center of motion (RCM) constraint is critical for sa

model-releasesarxiv-cs-ro
2 Jul 2026
Model Releases

ActivityNarrated: An Open-Ended Narrative Paradigm for Wearable Human Activity Understanding

DGX agent

arXiv:2604.00767v2 Announce Type: replace Abstract: Wearable human activity recognition (HAR) has made steady progress, yet much of this progress remains grounded in fixed-window, closed-set classific

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

AD-MPCC: Adaptive Differentiable Model Predictive Contouring Control for Autonomous Racing

DGX agent

arXiv:2607.00141v1 Announce Type: new Abstract: This paper presents Adaptive Differentiable Model Predictive Contouring Control (AD-MPCC), a framework for autonomous racing that integrates differentia

model-releasesarxiv-cs-ro
2 Jul 2026
Model Releases

Adversarial Pragmatics for AI Safety Evaluation: A Benchmark for Instruction Conflict, Embedded Commands, and Policy Ambiguity

DGX agent

arXiv:2607.01153v1 Announce Type: cross Abstract: Safety evaluations for language models increasingly depend on judgments about ambiguous natural-language behaviour: whether a model has followed an in

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AFFMAE: Scalable Vision Pre-Training for High-Resolution Microscopy Segmentation on Desktop Hardware

DGX agent

arXiv:2602.16249v2 Announce Type: replace Abstract: Self-supervised pretraining has transformed computer vision by enabling data-efficient fine-tuning, yet high-resolution pretraining typically requir

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

AGC-Bench: Measuring Artificial General Creativity

DGX agent

arXiv:2607.01152v1 Announce Type: new Abstract: Creativity research has debated whether creativity is domain-specific (e.g., visual, writing, science), and if it is psychometrically separable from gen

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

AGE: Adaptive-masking for Graph Embedding in Graph Retrieval-Augmented Generation

DGX agent

arXiv:2607.00052v1 Announce Type: cross Abstract: GraphRAG is an extension of retrieval-augmented generation (RAG) that supports large language models (LLMs) by referring to graph-structured data as e

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AGI Maze as a Benchmark Framework for World-Modeling Agents

DGX agent

arXiv:2607.00627v1 Announce Type: new Abstract: Large language models (LLMs) are powerful pattern-completion systems, but their default operating mode - predicting the next token from a static context

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AlgoBench: Benchmarking Algorithmic Adaptation in Code Generation

DGX agent

arXiv:2607.00062v1 Announce Type: cross Abstract: High pass rates on established programming benchmarks such as HumanEval and LiveCodeBench do not always show whether a model can reason about algorith

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Amortized Maximum Inner Product Search with Learned Support Functions

DGX agent

arXiv:2603.08001v3 Announce Type: replace Abstract: Maximum inner product search (MIPS) is a crucial subroutine in machine learning, requiring the identification of a vector taken within a database (t

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

An LLM-Based Framework for Intent-Driven Network Topology Design

DGX agent

arXiv:2607.00292v1 Announce Type: cross Abstract: Designing deployable and resilient network topologies from natural language requirements remains a challenging problem in network automation. This wor

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AnF-DiffPET: Anatomy- and Frequency-Guided Diffusion for PET/CT Denoising

DGX agent

arXiv:2607.00509v1 Announce Type: new Abstract: Positron emission tomography (PET) provides essential functional information for disease assessment, however reducing injected activity or acquisition t

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Another fascinating paper on LLM Judges. (bookmark it) It's from Amazon, and they show that if you run panels of LLM judges, averaging their…

DGX agent

Another fascinating paper on LLM Judges. (bookmark it) It's from Amazon, and they show that if you run panels of LLM judges, averaging their scores is a trap. 'Overall, we establish that robust aggreg

model-releasesdair-ai--x
2 Jul 2026
Model Releases

Are Performance-Optimization Benchmarks Reliably Measuring Coding Agents?

DGX agent

arXiv:2607.01211v1 Announce Type: cross Abstract: Repository-level performance-optimization benchmarks such as GSO, SWE-Perf and SWE-fficiency evaluate coding agents by applying patches to real reposi

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Artifacts in Claude Code have been life changing. Excited to expand to Pro and Max!

DGX agent

Artifacts in Claude Code have been life changing. Excited to expand to Pro and Max! Artifacts in Claude Code are now also available on Pro and Max plans. Ask for an artifact, Claude writes the code, p

model-releasesboris-cherny--x
2 Jul 2026
Model Releases

ATM: CID-Brokered Pre-Write Admission for Multi-Agent Code Co-Synthesis

DGX agent

arXiv:2607.00041v1 Announce Type: cross Abstract: Multi-agent LLM systems can decompose software-engineering work into planning, generation, validation, and repair, but a narrower systems problem rema

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Auditing Forgetting in Limited Memory Language Models

DGX agent

arXiv:2607.00605v1 Announce Type: cross Abstract: Limited Memory Language Models (LMLMs) externalize factual knowledge to a database to enable deletion-based unlearning without retraining. Existing ev

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AutoMem: Automated Learning of Memory as a Cognitive Skill

DGX agent

arXiv:2607.01224v1 Announce Type: new Abstract: Memory expertise is a learned skill: knowing what to encode, when to retrieve, and how to organize knowledge--a capacity known in cognitive science as m

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

// AutoMem // I quite like this idea of metamemory. (bookmark it) This new research from Stanford treats agent's memory management as a trai…

DGX agent

// AutoMem // I quite like this idea of metamemory. (bookmark it) This new research from Stanford treats agent's memory management as a trainable skill instead of a fixed module. The model decides wha

model-releasesdair-ai--x
2 Jul 2026
Model Releases

Autonomous Scientific Discovery via Iterative Meta-Reflection

DGX agent

arXiv:2607.01131v1 Announce Type: cross Abstract: Autonomous scientific discovery systems offer the potential to accelerate research by automating the process of hypothesis generation and validation.

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AV-SyncBench: Decoupled Benchmarking of Temporal and Semantic Audio-Visual Synchronization

DGX agent

arXiv:2607.00726v1 Announce Type: new Abstract: Audio-visual feature extraction is a fundamental component of multimodal understanding and generation tasks. However, existing evaluation protocols for

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

BaseRT: Best-in-Class LLM Inference on Apple Silicon via Native Metal

DGX agent

arXiv:2607.00501v1 Announce Type: cross Abstract: We present BaseRT, a native Metal inference runtime for large language models (LLMs) on Apple Silicon, and report the highest inference throughput on

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Bayesian Uncertainty Propagation for Agentic RAG Pipelines: A Proof-of-Concept Study on Multi-Hop Question Answering

DGX agent

arXiv:2607.00972v1 Announce Type: new Abstract: Trustworthy deployment of Agentic Retrieval-Augmented Generation (RAG) systems requires mechanisms for estimating when multi-stage reasoning pipelines m

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Benchmarking Frontier LLMs on Arabic Cultural and Sociolinguistic Knowledge: A Cross-Evaluation Framework with Human SME Ground Truth

DGX agent

arXiv:2607.00139v1 Announce Type: new Abstract: The cost of human expert evaluation is a principal bottleneck to deploying language models in specialized, high-stakes domains. This is particularly acu

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

Beyond Activation Alignment:The Alignment-Diversity Tradeoff in Task-Aware LLM Quantization

DGX agent

arXiv:2607.00908v1 Announce Type: new Abstract: Mixed-precision quantization (MPQ) has become a key technique for deploying large language models under stringent memory and compute constraints. We fir

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

Beyond Document Grounding: Span-Level Hallucination Detection over Code, Tool Output, and Documents

DGX agent

arXiv:2607.00895v1 Announce Type: new Abstract: Hallucination detection for retrieval-augmented generation (RAG) is usually evaluated on natural-language document evidence. However, grounded generatio

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

big week at langchain, with a lot of launches: 1/ OpenWiki - auto generate a wiki of a github repo 2/ two different voice agent tutorials 3/…

DGX agent

big week at langchain, with a lot of launches: 1/ OpenWiki - auto generate a wiki of a github repo 2/ two different voice agent tutorials 3/ Harbor integration and tutorial for long running, stateful

model-releasesharrison-chase--x
2 Jul 2026
Model Releases

Bridgewater just published numbers that should make every frontier lab nervous. The world's largest hedge fund tested Gemini, Claude, and GP…

DGX agent

Bridgewater just published numbers that should make every frontier lab nervous. The world's largest hedge fund tested Gemini, Claude, and GPT on six document filtering tasks its investors do every day

model-releasesclem-delangue--x
2 Jul 2026
Model Releases

But yikes does Fable write text that sounds like a parody of a Claude model on overdrive.

DGX agent

This post critiques Fable AI's text generation style, suggesting it produces overly verbose or exaggerated outputs that parody Claude's characteristic writing patterns taken to an extreme. The comment

model-releasesethan-mollick--x
2 Jul 2026
Model Releases

Can Agents Generalize to the Open World? Unveiling the Fragility of Static Training in Tool Use

DGX agent

arXiv:2607.01084v1 Announce Type: new Abstract: While Large Language Model (LLM) agents demonstrate proficiency in static benchmarks, their deployment in real-world scenarios is hindered by the dynami

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Clinician-Level Agreement Without Clinical Caution: LLM Evaluator Limits in Medical AI Benchmarking

DGX agent

arXiv:2607.01103v1 Announce Type: new Abstract: Open-response evaluation provides stronger clinical validity than multiple-choice benchmarks but creates a scoring bottleneck that motivates automated L

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

Comparing Large Language Models on Scrum Certification-Style Questions: Accuracy, Stability, and Error Patterns

DGX agent

arXiv:2607.00048v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in exam- and certification-style question answering tasks, where their ability to retrieve, interpr

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Computer vision-based neural networks for radioisotope identification in urban environments

DGX agent

arXiv:2607.00270v1 Announce Type: cross Abstract: Algorithm development for radioisotope identification in mobile urban search scenarios face significant challenges from non-uniform backgrounds, momen

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

Continual learning is probably the biggest barrier to explosive AI adoption (& may have big implications for recursive self-improvement as w…

DGX agent

Continual learning is probably the biggest barrier to explosive AI adoption (& may have big implications for recursive self-improvement as well) As long as you deal with amnesiac models that require h

model-releasesethan-mollick--x
2 Jul 2026
Model Releases

CoT-X: An Adaptive Framework for Cross-Model Chain-of-Thought Transfer and Optimization

DGX agent

arXiv:2511.05747v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning enhances the problem-solving ability of large language models (LLMs) but leads to substantial inference overhead, l

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

CPDDNet: Color-Polarization Denoising and Demosaicking Network

DGX agent

arXiv:2607.01100v1 Announce Type: new Abstract: Color-polarization imaging using a color-polarization filter array (CPFA) sensor captures both texture (color intensity) and physical (polarization) inf

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Creating Impactful Autonomous Driving Datasets: A Strategic Guide from Research Gap to Benchmark

DGX agent

arXiv:2607.00710v1 Announce Type: cross Abstract: Well-designed autonomous driving datasets have fundamentally shaped research progress, yet existing literature primarily describes what datasets conta

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

DeepSeek effect https://x.com/maximelabonne/status/2070867418377818542

DGX agent

DeepSeek effect https://x.com/maximelabonne/status/2070867418377818542 Fun surprise: DeepSeek used my open-perfectblend dataset to train their new DSpark drafter Time to promote it again! It's an open

model-releasesclem-delangue--x
2 Jul 2026
Model Releases

DiscoLoop: Looping Discrete Embeddings and Continuous Hidden States for Multi-hop Reasoning

DGX agent

arXiv:2607.00341v1 Announce Type: cross Abstract: Large language models achieve strong performance on many reasoning tasks when allowed to externalize intermediate steps as Chain-of-Thought (CoT). How

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Distributionally Robust Linear Regression With Block Lewis Weights

DGX agent

arXiv:2607.00252v1 Announce Type: new Abstract: We present an algorithm for the group distributionally robust (GDR) least squares problem. Given m groups, a parameter vector in R^d, and stacked design

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

Does Your ViT Still Need U-Net for Segmentation?

DGX agent

arXiv:2607.00223v1 Announce Type: new Abstract: Medical image segmentation is dominated by U-Net-style encoder-decoder architectures. Vision Transformers (ViTs) overcome the limited receptive field of

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

DriveVA: Video Action Models are Zero-Shot Drivers

DGX agent

arXiv:2604.04198v2 Announce Type: replace Abstract: Generalization is a central challenge in autonomous driving, as real-world deployment requires robust performance under unseen scenarios, sensor dom

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

DriveVer: Lightweight Trajectory Evaluator as Test-Time Verifier for Autonomous Driving

DGX agent

arXiv:2607.00399v1 Announce Type: new Abstract: End-to-end autonomous driving models often encounter performance bottlenecks, as training-time scaling leads to high computational costs and diminishing

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

DroneFINE: Domain-Aware Parameter-Efficient Fine-Tuning of Vision-Language Detectors for Drone Images

DGX agent

arXiv:2607.00338v1 Announce Type: new Abstract: Object detection for Unmanned Aerial Vehicles (UAVs) working in open and dynamic environments is a highly challenging task. While Vision-Language Models

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Dual-Confidence Contrastive Decoding for Retrieval-Augmented Generation

DGX agent

arXiv:2607.00570v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) increasingly requires models to answer questions from multiple retrieved documents, where only some sources are rel

model-releasesarxiv-cs-cl
2 Jul 2026
← Previous
1…135136137138139…471
Next →