AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,585 results
Model Releases

“That was fun, I’ll do it again when I get out,” said the 22-year-old Gambian violent offender after he stabbed a 55-year-old Italian man tw…

DGX agent

“That was fun, I’ll do it again when I get out,” said the 22-year-old Gambian violent offender after he stabbed a 55-year-old Italian man twenty times. And he will be released again. Perhaps to succee

model-releaseselon-musk--x
5 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

There is a massive demand for processing files *in the agent loop*. The number of users submitting agent queries with file attachments is ex…

DGX agent

There is a massive demand for processing files *in the agent loop*. The number of users submitting agent queries with file attachments is exponentially increasing over time. Our mission is to make Lit

model-releasesjerry-liu--x
5 Jul 2026
Model Releases

“What has really happened is VC built the wrong AI, you’ll pay a high price for this error with your pension, and nobody will go to jail.”

DGX agent

“What has really happened is VC built the wrong AI, you’ll pay a high price for this error with your pension, and nobody will go to jail.” By far the shrewdest and most entertaining analyst of the AI

model-releasesgary-marcus--x
5 Jul 2026
Model Releases

Wordle 1,841 4/6 ⬛⬛🟨⬛⬛ ⬛🟩⬛⬛⬛ ⬛🟩⬛⬛🟩 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle game result (puzzle #1,841) solved in 4 attempts, showing the letter position feedback for each guess using the standard color-coded system (gray for incorrect letters, ye

model-releasesanthropic--x
5 Jul 2026
Model Releases

As America turns 250, we put together 250 open AI milestones from the US: open models, datasets, demos, papers, and tools that helped shape …

DGX agent

As America turns 250, we put together 250 open AI milestones from the US: open models, datasets, demos, papers, and tools that helped shape the field. They go from attention is all you need, pytorch,

model-releasesclem-delangue--x
4 Jul 2026
Model Releases

Better Models: Worse Tools

DGX agent

Better Models: Worse Tools Armin reports on a weird problem he ran into while hacking on Pi: The short version is that newer Claude models sometimes call Pi’s edit tool with extra, invented fields in

model-releasessimon-willison
4 Jul 2026
Model Releases

Hey, Fable: 'No like 'AAA' in what everyone online thinks AAA is, you know what I mean.' This was pretty funny, there are lootboxes, EULAs, …

DGX agent

Hey, Fable: 'No like 'AAA' in what everyone online thinks AAA is, you know what I mean.' This was pretty funny, there are lootboxes, EULAs, achievements, useless graphical settings, elaborate boot scr

model-releasesethan-mollick--x
4 Jul 2026
Model Releases

i often think about the irony of how 'tools for thought' people spent like a decade making cool pretty demos with canvases and then got comp…

DGX agent

i often think about the irony of how 'tools for thought' people spent like a decade making cool pretty demos with canvases and then got completely mogged by low contrast poorly designed CLIs just winn

model-releasesswyx--x
4 Jul 2026
Model Releases

Indeed: if we actually had AGI we would not need forward deployed engineers.

DGX agent

Indeed: if we actually had AGI we would not need forward deployed engineers. Two simultaneous AI narratives right now: 1) You can now do the work of 20 people, learn anything and create anything with

model-releasesgary-marcus--x
4 Jul 2026
Model Releases

introducing tinyrouter i reverse engineered the routing architecture behind Skana AI's Fugu and built replication for open frontier models. …

DGX agent

introducing tinyrouter i reverse engineered the routing architecture behind Skana AI's Fugu and built replication for open frontier models. it's a tiny ~10K parameter LLM router that learns which mode

model-releasesclem-delangue--x
4 Jul 2026
Model Releases

spending the last week at @aidotengineer was awesome. too many great convos to cover them all, but jotted down some things that stood out: -…

DGX agent

spending the last week at @aidotengineer was awesome. too many great convos to cover them all, but jotted down some things that stood out: - Lots of discussion around open source models. I spoke with

model-releasesyohei-nakajima--x
4 Jul 2026
Model Releases

The fanfiction community is at war with AI — and itself

DGX agent

Over the past week, a new fanworks movement has kicked off, with the aim to root out authors using generative AI. But the detection methods being implemented are questionable, and any fanfic writer co

model-releasesthe-verge-ai
4 Jul 2026
Model Releases

The July 4th weekend All-In @theallinpod turned into a long argument about who owns the intelligence layer. The besties think enterprises ju…

DGX agent

The July 4th weekend All-In @theallinpod turned into a long argument about who owns the intelligence layer. The besties think enterprises just woke up to a trap they had been walking into, here's how

model-releasesclem-delangue--x
4 Jul 2026
Model Releases

We've created a comprehensive Retrieval Harness for modern agentic retrieval in 2026. The harness provides a persistent data pipeline that c…

DGX agent

We've created a comprehensive Retrieval Harness for modern agentic retrieval in 2026. The harness provides a persistent data pipeline that can connect to a data source, index and update a large knowle

model-releasesjerry-liu--x
4 Jul 2026
Model Releases

A behind-the-scenes look at Midjourney’s medical scanner leaves many questions unanswered

DGX agent

Midjourney has shown more of its futuristic medical scanner. It still hasn't shown much proof it works. The AI startup, best known for generating images, released a behind-the-scenes video of its dunk

model-releasesthe-verge-ai
3 Jul 2026
Model Releases

A global predicted-fMRI drive signal from TRIBE does not predict YouTube replay heatmaps

DGX agent

arXiv:2607.01400v1 Announce Type: cross Abstract: Deep multimodal brain-encoding models now predict fMRI responses to naturalistic video with high accuracy. Whether their predicted neural signals also

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

A Memory Efficient Unified Algorithm for Online Learning of Linear Dynamical Systems

DGX agent

arXiv:2607.02050v1 Announce Type: new Abstract: Motivated by the challenge of stabilizing a general unknown linear dynamical system (LDS) from observations, we study the natural prerequisite of online

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

A rubric-based controlled comparison of frontier language models on expert-authored clinical reasoning tasks

DGX agent

arXiv:2607.02175v1 Announce Type: new Abstract: Multiple-choice medical benchmarks are increasingly saturated, and recent rubric-based evaluations such as HealthBench have shown that open-ended clinic

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

A-TMA: Decoupling State-Aware Memory Failures in Long-Term Agent Memory

DGX agent

arXiv:2607.01935v1 Announce Type: new Abstract: Long term memory lets LLM agents act as persistent assistants, but user facts change. A useful memory system must know what is true now, what used to be

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

A^{2}utoLPBench: An Auto-Generated, Agent-Friendly LP Benchmark via Inverse-KKT Construction

DGX agent

arXiv:2607.02141v1 Announce Type: new Abstract: Most LP-from-text benchmarks are static datasets of word problems written and labeled by hand. Once such a dataset is released, its size is fixed, its d

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

AbsoluteDegradation: A Physics-Inspired Synthetic Film-Degradation Pipeline and Archival Film Restoration Benchmark

DGX agent

arXiv:2607.02131v1 Announce Type: cross Abstract: Restoring archival film remains a fundamentally challenging problem due to the absence of paired training data and the lack of standardized evaluation

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence

DGX agent

arXiv:2604.08169v2 Announce Type: replace Abstract: Alignment in LLMs is more brittle than commonly assumed: misalignment can be induced by adversarial prompts, benign fine-tuning, emergent misalignme

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Adaptive Batch Sizes Using Non-Euclidean Gradient Noise Scales for Stochastic Sign and Spectral Descent

DGX agent

arXiv:2602.03001v2 Announce Type: replace-cross Abstract: To maximize hardware utilization, modern machine learning systems typically employ large constant or manually tuned batch size schedules, rely

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Adoption and Impact of Command-Line AI Coding Agents: A Study of Microsoft's Early 2026 Rollout of Claude Code and GitHub Copilot CLI

DGX agent

arXiv:2607.01418v1 Announce Type: cross Abstract: Organizations rolling out agentic command line tools like Anthropic's Claude Code and GitHub's Copilot CLI need to know who will try them, who will ke

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Agent4cs: A Multi-agent System for Code Summarization in Large Hierarchical Codebases

DGX agent

arXiv:2607.01425v1 Announce Type: new Abstract: Understanding large, complex codebases, especially those with obfuscated structures and incomplete documentation, remains a significant challenge. Exist

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

AgenticDataBench: A Comprehensive Benchmark for Data Agents

DGX agent

arXiv:2607.01647v1 Announce Type: cross Abstract: Data science aims to derive actionable insights from heterogeneous raw data, unlocking the value of the massive amounts of data generated in modern so

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

AgenticRAGTracer: A Hop-Aware Benchmark for Diagnosing Multi-Step Retrieval Reasoning in Agentic RAG

DGX agent

arXiv:2602.19127v2 Announce Type: replace Abstract: With the rapid advancement of agent-based methods in recent years, Agentic RAG has undoubtedly become an important research direction. Multi-hop rea

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

AgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM Agents

DGX agent

arXiv:2607.02255v1 Announce Type: new Abstract: Memory for a long-horizon LLM agent is a contract about what each future decision is allowed to see. The simplest contract appends past observations, to

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

AIriskEval-edu: New Dataset for Risk Assessment in AI-mediated K-12 Educational Explanations

DGX agent

arXiv:2607.01934v1 Announce Type: cross Abstract: This work introduces AIriskEval-edu-db2, a new dataset designed to train and evaluate auditors based on LLMs for an explainable pedagogical risk asses

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

An Isotropic Approach to Efficient Uncertainty Quantification with Gradient Norms

DGX agent

arXiv:2603.29466v2 Announce Type: replace-cross Abstract: Existing methods for quantifying predictive uncertainty in neural networks are either computationally intractable for large language models or

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Anthropic wants to develop its own drugs

DGX agent

At the event 'The Briefing: AI for Science' earlier this week, Anthropic announced Claude Science, a new 'AI workbench for scientists' that pulls fragmented tools and datasets into one environment, an

model-releasesthe-verge-ai
3 Jul 2026
Model Releases

AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models

DGX agent

arXiv:2607.02269v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated immense promise in Spatio-Temporal Video Grounding (STVG). However, current evaluation protocols are l

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Ask the Right Comparison:Bias-Aware Bayesian Active Top-k Ranking with LLM Judges

DGX agent

arXiv:2607.02104v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as cheap, scalable judges that compare candidate outputs pairwise -- to rank responses, select models

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Assessing VLM Reliability for Medical Image Quality Evaluation Under Corruption and Bias

DGX agent

arXiv:2607.01973v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly applied in medical tasks such as pathology description, report generation, and visual question answerin

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Automated grading of Linux/bash examinations using large language models: a four-level cognitive taxonomy approach

DGX agent

arXiv:2607.02432v1 Announce Type: new Abstract: Scalable and reliable grading of command-line examinations remains a challenge in computing education, where rising enrolments make manual marking diffi

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

BALF: Budgeted Activation-Aware Low-Rank Factorization for Fine-Tuning-Free Model Compression

DGX agent

arXiv:2509.25136v3 Announce Type: replace Abstract: Activation-aware low-rank factorization techniques yield strong compression results but are generally confined to linear layers, while existing whit

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Bayesian Sparse Low-Rank Adaptation for Large Language Model Uncertainty Estimation

DGX agent

arXiv:2607.02182v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit remarkable reasoning capabilities, but their task-specific fine-tuning is notoriously plagued by overconfidence,

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Benchmarking Federated Learning and Knowledge Distillation for Point Cloud Classification

DGX agent

arXiv:2607.01272v1 Announce Type: cross Abstract: Deploying 3D point cloud analysis in privacy-sensitive, resource-constrained settings faces two barriers: data cannot be centralized, and models must

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Beyond Pixel Diffs: Benchmarking Image Change Captioning for Web UI Visual Regression Testing

DGX agent

arXiv:2607.01728v1 Announce Type: cross Abstract: Visual regression testing (VRT) is a standard quality assurance step in modern software release pipelines. On every change, it re-renders user interfa

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Beyond Skepticism: Evaluating LLMs Pedagogical Intent Reasoning with the Adaptive Pedagogical Vigilance Framework

DGX agent

arXiv:2607.01581v1 Announce Type: new Abstract: The capacity of Large Language Models (LLMs) to reason about pedagogical intent within instructional communication remains underexplored, particularly i

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Black-Box Inference of LLM Architectural Properties with Restrictive API Access

DGX agent

arXiv:2607.01313v1 Announce Type: cross Abstract: In practice, most commercial LLM providers do not publicly release details of underlying LLM architectures. However, prior work has shown that given l

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Born Discrete, Made Smooth: Variational Formulation of Shallow Neural Networks

DGX agent

arXiv:2607.02003v1 Announce Type: cross Abstract: Although neural networks are remarkably effective, their underlying optimization principles remain theoretically elusive, often characterized by non-c

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Boundary-Aware Quantization: Finite-Scale Decision Geometry of Neural Classifiers

DGX agent

arXiv:2607.01478v1 Announce Type: cross Abstract: We measured quantization-induced decision-boundary changes using local logit-margin radii, first-order boundary displacement, normal variation, slice-

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

BOUNDARY_SYNC: Measuring Communication-Induced Representational Coupling in Multi-Agent LLM Systems

DGX agent

arXiv:2607.01600v1 Announce Type: cross Abstract: As large language models (LLMs) are deployed as communicating agents, does inter-agent communication cause outputs to converge? We introduce BOUNDARY_

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Breaking Safety at the Token Boundary: How BPE Tokenization Creates Exploitable Gaps in LLM Alignment

DGX agent

arXiv:2607.01239v1 Announce Type: cross Abstract: Character-level perturbations bypass safety alignment in modern LLMs despite leaving prompts human-readable. We identify and test a central structural

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

BRIDGE: Predicting Human Task Completion Time From Model Performance

DGX agent

arXiv:2602.07267v2 Announce Type: replace Abstract: Evaluating the real-world capabilities of AI systems requires grounding benchmark performance in human-interpretable measures of task difficulty. Ex

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Bringing Agentic Search to Earth Observation Data Discovery

DGX agent

arXiv:2607.02387v1 Announce Type: cross Abstract: NASA and its data centers hold thousands of geoscience datasets and tools like Worldview, Giovanni, the Science Discovery Engine, and Harmony. Finding

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

BuilderBench: The Building Blocks of Intelligent Agents

DGX agent

arXiv:2510.06288v4 Announce Type: replace Abstract: Today's AI models learn primarily through mimicry and refining, so it is not surprising that they struggle to solve problems beyond the limits set b

model-releasesarxiv-cs-ai
3 Jul 2026
← Previous
1…132133134135136…471
Next →