AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,323 results
Model Releases

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds

DGX agent

arXiv:2608.02636v1 Announce Type: cross Abstract: Self-evolving skill systems promise to improve agents by turning execution feedback into persistent skill updates without changing the underlying mode

model-releasesarxiv-cs-ai
5 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Reversing Arrows in Large Language Models

DGX agent

arXiv:2608.03512v1 Announce Type: new Abstract: Large language models (LLMs) have achieved strong performance on text-to-knowledge graph generation and related tasks. Nevertheless, it is still unclear

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Risky Business: Measuring The Faithfulness-Safety Tension

DGX agent

arXiv:2608.03745v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning offers a promising window into model monitoring. However, monitoring relies on faithfulness, i.e., the model output str

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Route-Align-Verify for Functional Correctness in Code Generation

DGX agent

arXiv:2608.03341v1 Announce Type: cross Abstract: Large language models (LLMs) have substantially improved code generation, yet achieving strong functional correctness remains difficult, especially fo

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Rubrics as Privileged Information for Open-Ended Generation

DGX agent

arXiv:2608.02948v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD), where a single model acts as both student and teacher with different contexts, has shown promise in verifiable dom

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

S^3: Improving Agent Safety through Multi-Stage Defense

DGX agent

arXiv:2608.02683v1 Announce Type: cross Abstract: Large Language Model (LLM) agents rely on multi-stage agentic workflows, with stages such as memory, planning, and tool execution, to accomplish compl

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

SAKI: Score-Aware Low-Rank Key Indexing for Long-Context KV Retrieval

DGX agent

arXiv:2608.03228v1 Announce Type: new Abstract: Existing low rank KV cache methods preserve either model weights or key variance, neither of which directly reflects the attention scores used during in

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Scaling agentic AI: How UiPath built its high-performance GPU platform on AI Hypercomputer

DGX agent

As a market leader in enterprise agentic automation and business orchestration, UiPath is helping to pioneer an industry shift toward agentic AI. With it, the company is deploying autonomous agents to

model-releasesgoogle-cloud-ai
5 Aug 2026
Model Releases

Scenema Audio Comes to ComfyUI, Runs on 8GB VRAM

DGX agent

Hey everyone! Scenema Audio is now a native ComfyUI custom node. Same model that powers scenema.ai now quantized so it fits on 8GB VRAM. When we first released it a few months ago as an API and Docker

model-releasesr-localllama
5 Aug 2026
Model Releases

SciRet: A Compute-Aware Empirical Study of Retrieval and Reranking for Scientific RAG

DGX agent

arXiv:2608.03860v1 Announce Type: cross Abstract: We introduce SciRet, a compute-aware empirical study of retrieval-augmented generation for scientific question answering over CORD-19. Rather than pro

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Screenshots or Tools? Eliciting Tool Use and Managing Multimodal Context in Hybrid GUI-MCP Computer-Use Agents

DGX agent

arXiv:2608.03327v1 Announce Type: new Abstract: Hybrid computer-use agents can act through screenshots or call text tools. We find that having a tool available does not settle which way the effect goe

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

SeaSlides: Semantic Abstraction Layer for Agentic Slide Generation

DGX agent

arXiv:2608.03298v1 Announce Type: new Abstract: Agentic presentation generation must preserve source content, maintain coherent visual design, render specialized objects, and produce usable artifacts.

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Security-First Evaluation of Text-to-Terraform: Benchmarking LLMs and SLMs for Secure IaC Generation

DGX agent

arXiv:2608.02672v1 Announce Type: cross Abstract: Cloud misconfiguration remains a leading cause of security incidents, yet whether LLMs and SLMs can generate security-compliant Infrastructure-as-Code

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Sensitivity, Causality, and Repair Dissociate: A Layer-Wise Analysis of Perturbation Robustness and Its Scaling

DGX agent

arXiv:2608.03842v1 Announce Type: new Abstract: When a language model fails on surface-perturbed input (typos, OCR noise, homophones), 'which layer is responsible' has three natural operationalization

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

SeqLLM: Augmenting LLMs with Behavioral-Sequence Modeling for High-Stakes Decisions at WeChat Pay

DGX agent

arXiv:2608.03063v1 Announce Type: new Abstract: Merchant risk control at large payment platforms screens tens of millions of merchants daily, where false positives harm legitimate merchants and false

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMs

DGX agent

arXiv:2608.03573v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL) exhibit fundamentally different behaviors in enhancing multi-task reasoning for large langu

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Shorter Reasoning, Earlier Answers? An Evaluation of Reasoning Interfaces

DGX agent

arXiv:2608.03401v1 Announce Type: cross Abstract: Large language models often reason at length before answering, increasing cost and latency. Prompts and trained settings can shorten this reasoning, b

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA

DGX agent

arXiv:2509.25459v2 Announce Type: replace Abstract: Large Language Models (LLMs) show promise in generating long-form scientific explanations that synthesize evidence and connect multiple factors. How

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Single Canonical Prompts Underestimate LLM Safety's Surface-Form Sensitivity

DGX agent

arXiv:2608.02665v1 Announce Type: cross Abstract: A benchmark score is a measurement instrument, yet most benchmarks read each item at a single canonical surface form. We ask whether that reading is f

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

SITUATION DETECTED: Prime Intellect is releasing Prime Agent, a self-improving harness for coding and long-running autonomous tasks. The tea…

DGX agent

SITUATION DETECTED: Prime Intellect is releasing Prime Agent, a self-improving harness for coding and long-running autonomous tasks. The team reports 95.5% on ARC-AGI-3, above the human baseline, and

model-releasesyohei-nakajima--x
5 Aug 2026
Model Releases

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. C…

DGX agent

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. ContinualSkillBench covers five domains, each with 100 interc

model-releasesdair-ai--x
5 Aug 2026
Model Releases

SlimVLM: Sensitivity-aware Dynamic Structured Pruning with Adaptive Visual Token Selection for Efficient Vision-Language Models

DGX agent

arXiv:2608.03580v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have demonstrated remarkable performance in processing and understanding both text and images, their large parameter

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

SocietyBench: Forecasting Counterfactual Social-World Evolution

DGX agent

arXiv:2608.04009v1 Announce Type: new Abstract: Large language models (LLMs), and the agents built on top of them, are now benchmarked heavily on whether they can finish a task -- fix a bug, drive a b

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

SpatialCLI: Learning to Reason With Spatial Tools, Then Without Them

DGX agent

arXiv:2607.27703v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly used in embodied agents to interpret visual inputs, reason about spatial relationships, and make task

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Sphere Retraction Normalizations

DGX agent

arXiv:2608.02668v1 Announce Type: cross Abstract: Residual connections are the de facto mechanism for training deep neural networks stably. Geodesic Normalization (GeoNorm) recasts them on a Riemannia

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Stable Diffusion might actually be remembered in the history books, and I don’t think that’s an overstatement

DGX agent

Hear me out before you roll your eyes. We tend to only recognize turning points in hindsight. Nobody in 1993 thought the Mosaic browser would be a history book moment, but the web is. I think Stable D

model-releasesr-stablediffusion
5 Aug 2026
Model Releases

StereoVGGT: A Training-Free Visual Geometry Transformer for Stereo Vision

DGX agent

arXiv:2603.29368v2 Announce Type: replace Abstract: Driven by the advancement of 3D devices, stereo vision tasks including stereo matching and stereo conversion have emerged as a critical research fro

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

STREAM-VAE: Dual-Path Routing for Slow and Fast Dynamics in Vehicle Telemetry Anomaly Detection

DGX agent

arXiv:2511.15339v3 Announce Type: replace-cross Abstract: Automotive telemetry data exhibits slow drifts and fast spikes, often within the same sequence, making reliable anomaly detection challenging.

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Stuck on 'A': Diagnosing and Repairing Interface Injury in Attention-to-KDA Linearization of a 0.6B Language Model

DGX agent

arXiv:2608.02689v1 Announce Type: new Abstract: We convert 21 of 28 full-attention layers of Qwen3-0.6B-Base into KDA (Kimi Delta Attention) linear-attention layers on a single consumer-grade GPU budg

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Sulphur 3 is looking for funding

DGX agent

Hello, I'm the guy who made Sulphur 2. With the recent release of a certain video model, we are looking to mobilize and train Sulphur 3 on this new model. We are targeting $10,000 USD. This certain ne

model-releasesr-stablediffusion
5 Aug 2026
Model Releases

SUV: Future Scene Understanding as Video Generation for End-to-End Driving

DGX agent

arXiv:2608.03084v1 Announce Type: new Abstract: End-to-end driving requires a coherent understanding of future scenes, yet existing methods model these scenes using task-specific heads and output form

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

TACT: Taxonomy-Aligned Post-Training for Pedagogically Adaptive English Tutoring

DGX agent

arXiv:2608.03952v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to provide conversational practice for English-as-a-second-language (ESL) learners. Effective ESL tut

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Taming the Implicit: Dual-Channel Risk-Aware Reinforcement Fine-Tuning for Continual Multimodal Post-Training

DGX agent

arXiv:2608.03660v1 Announce Type: new Abstract: Reinforcement fine-tuning (RFT) is widely believed to inherently resist catastrophic forgetting in continual post-training of multimodal large language

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

TARL: Transaction-Aware Reliable Ledgers for Executable Memory Management in Long-Term Agents

DGX agent

arXiv:2608.03699v1 Announce Type: new Abstract: Persistent memory helps long-term agents retain knowledge, yet a single update error can repeatedly distort future retrieval and reasoning. Most existin

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Task-Oriented Candidate-Latent Feedback for Coarse-to-Fine Sensing in Distributed OFDM-ISAC Networks

DGX agent

arXiv:2608.03319v1 Announce Type: cross Abstract: Future integrated sensing and communication (ISAC) architectures separate the sensing entity (SE) that acquires measurements from the sensing function

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Test-Time Augmentation for Tabular-to-Image Classifiers under Distribution Shifts

DGX agent

arXiv:2608.03557v1 Announce Type: new Abstract: Tabular-to-image methods that convert tabular data into visual representations have emerged as a novel paradigm for leveraging the high performance of d

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

The Google Bug Hunters Team admitted to me that they cannot fundamentally patch prompt engineering bypasses in Gemini

DGX agent

Hello everyone Yesterday I gave a report on Gemini bugs and the techniques I learned on Gemini so far with the Engineering Prompt and interestingly today I got a very interesting and controversial ans

model-releasesr-chatgpt
5 Aug 2026
Model Releases

The Ignition Is Real, and It Lives at the Readout: Latent composition, difficulty-clocked ignition, and the interface-constituted commit in a recurrent-depth reasoner

DGX agent

arXiv:2608.03263v1 Announce Type: cross Abstract: We test whether the 'compositional ignition' reported in latent-reasoning models is real computation, an instrument artifact, or inherited from verbal

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

The production of meaning in the processing of natural language

DGX agent

arXiv:2603.20381v2 Announce Type: replace-cross Abstract: Understanding the fundamental mechanisms governing the production of meaning in the processing of natural language is critical for designing s

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Thinking of buying more DRAM right now...

DGX agent

So I'm looking at https://huggingface.co/unsloth/DeepSeek-V4-Flash-0731-GGUF and I realize my 128GB of DRAM just isn't cutting it for this (incredibly powerful) model. If only I had another 64GB, I th

model-releasesr-localllama
5 Aug 2026
Model Releases

Third-party cyber evaluations involving OpenAI models

DGX agent

Third-party cyber evaluations involving OpenAI models And another one. I had to create a accidental-cyberattacks tag to keep track of them all! This post from OpenAI covers both the UK AI Safety Insti

model-releasessimon-willison
5 Aug 2026
Model Releases

TimeRLM: Recursive Language Models Enable Precise Anomaly Localization in Long-Context Time-Series

DGX agent

arXiv:2608.03391v1 Announce Type: new Abstract: Precise anomaly localization over long-context time series is a crucial task in monitoring applications across clinical care, industrial operations, fin

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

TraceCompiler: Skill-Guided Mining and Compilation of LLM Agent Traces into Mostly Deterministic Workflows

DGX agent

arXiv:2608.02680v1 Announce Type: cross Abstract: Tool-using language-model agents repeatedly rediscover procedures they have already executed, producing traces that mix reusable structure with retrie

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Trajectory inference via Acceleration Matching

DGX agent

arXiv:2608.03916v1 Announce Type: new Abstract: Trajectory inference is a fundamental problem in many scientific domains: given a collection of unpaired snapshots of observations at discrete time poin

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Try npm i -g cline on Cline!👀

DGX agent

Try npm i -g cline on Cline!👀 Qwen3.8-Max is now available in ClinePass, a subscription for ~5x discounted access. Use it on Cline CLI w/ $4.99 special promo: npm i -g cline This is currently the most

model-releasesqwen--x
5 Aug 2026
Model Releases

TumorBoard: Evidence-Grounded Multi-Agent Decision Support for Longitudinal Neuro-Oncology

DGX agent

arXiv:2608.03190v1 Announce Type: new Abstract: Neuro-oncology decisions require coordinated interpretation of serial MRI, pathology, molecular markers, treatment history, performance status, and evol

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Unequal Verdicts: Investigating Gender Bias in LLM-Based Fake News Detection

DGX agent

arXiv:2608.03627v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used for automated fact-checking, yet their susceptibility to gender bias in this context remains underexp

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

UniNav: A Unified World-Action Diffusion Model for Visual Navigation

DGX agent

arXiv:2608.03244v1 Announce Type: new Abstract: Image-goal visual navigation is a fundamental capability for embodied agents. Existing navigation policies efficiently predict waypoint trajectories but

model-releasesarxiv-cs-ai
5 Aug 2026
← Previous
1…4041424344…466
Next →