AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlog
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
Model Releases

Replit Free Mode, powered by @OpenAI GPT-5.6 Luna, helps you maximize making while minimizing token costs.

DGX agent

Replit introduced **Free Mode** on August 19, 2026, powered by OpenAI’s GPT‑5.6 Luna model. The feature is designed to let users “maximise making while minimizing token costs,” offering free access wi

model-releasesreplit--x
19 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Thinking in a Low-Resource Language: What SFT Builds, What RL Fixes, What Accuracy Cannot See

DGX agent

arXiv:2608.17744v1 Announce Type: new Abstract: Take three frontier mixture-of-experts models (Alibaba, OpenAI, NVIDIA; 3.6-4.0B active parameters each) and fine-tune them to reason in a low-resource

model-releasesarxiv-cs-cl
19 Aug 2026
Model Releases

When Personalization Becomes Bias: Structural and Discursive Religious Framing in AI-Generated Financial Advice

DGX agent

arXiv:2608.16909v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into financial advisory systems, yet their role in reproducing religious bias remains underex

model-releasesarxiv-cs-ai
19 Aug 2026
Local Ai

When to Plan, When to Polish: Noise Level as a Granularity Axis for Diffusion Language Models

DGX agent

arXiv:2606.21802v2 Announce Type: replace Abstract: Standard tokenwise diffusion LMs keep training corruption and inference commitment at token granularity throughout denoising. At high noise, this le

local-aiarxiv-cs-cl
19 Aug 2026
Research

A cross-modal generative model for incomplete and degraded prostate MRI with multicentre clinical validation

DGX agent

arXiv:2608.16233v1 Announce Type: cross Abstract: Missing or degraded sequences can limit prostate multiparametric MRI. We developed MSCNet, a sequence-conditioned cross-modal generative framework for

researcharxiv-cs-ai
18 Aug 2026
Research

A Deep Learning Model for Spatially Clustered Data via Differentiable Cluster Assignment

DGX agent

arXiv:2608.14968v1 Announce Type: cross Abstract: We consider nonparametric regression when the association between a response and its covariates changes across an unknown partition of a spatial domai

researcharxiv-cs-lg
18 Aug 2026
Model Releases

AeroCopilotBench: A Two-Tier Benchmark for Evaluating LLM Agents as Aviation Copilots in an Interactive Virtual Cockpit Environment

DGX agent

arXiv:2608.16349v1 Announce Type: new Abstract: Large language model (LLM) agents may assist flight crews with complex decisions and task execution, but existing aviation evaluations centered on stati

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

AeroGround: A Comprehensive Benchmark for Aerial-Ground Collaborative Reasoning

DGX agent

arXiv:2608.14721v1 Announce Type: new Abstract: Vision-language models (VLMs) have been widely employed in understanding and reasoning tasks for unmanned aerial vehicles (UAVs). Existing UAV benchmark

model-releasesarxiv-cs-cv
18 Aug 2026
Model Releases

Ask, Condition or Abstain: Reinforcement Learning for Missing-Premise Reasoning

DGX agent

arXiv:2608.16554v1 Announce Type: new Abstract: Answer-only reinforcement learning (RL) trains reasoning models to solve fully specified problems, but many realistic queries omit a premise needed for

model-releasesarxiv-cs-cl
18 Aug 2026
Model Releases

Auxiliary uncertainty signals for LLM-assisted systematic review screening: a benchmark across eight Cohen drug-class reviews

DGX agent

arXiv:2608.14551v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for title-abstract screening in systematic reviews, but their decisions lack calibrated uncertainty.

model-releasesarxiv-cs-cl
18 Aug 2026
Model Releases

Bye-bye, Bluebook? Automating Legal Drudgery With AI-Augmented Rule Following

DGX agent

arXiv:2505.02763v2 Announce Type: replace-cross Abstract: One of the central promises of legal AI is to automate drudgery -- the formal, repetitive tasks of lawyers' work that consume time without cal

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Can LLMs Reason Like Automated Theorem Provers for Rust Verification? VCoT-Bench: Evaluating via Verification Chain of Thought

DGX agent

arXiv:2603.18334v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) increasingly assist secure software development, their ability to meet the rigorous demands of Rust program ve

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Disentangling Pictorial Cue Understanding from Language Bias in VLMs via Depth Ordering Task

DGX agent

arXiv:2607.01503v2 Announce Type: replace Abstract: In this paper, we study depth perception of vision-language models (VLMs) to isolate the effects of pictorial depth cues and disentangle vision and

model-releasesarxiv-cs-cv
18 Aug 2026
Model Releases

Do LLM Agents Negotiate Rationally? A Mechanism-Design Framework for Verifiable Multi-Agent Interaction over A2A/MCP

DGX agent

arXiv:2608.14613v1 Announce Type: new Abstract: Modern LLM-agent frameworks increasingly interoperate through standards such as Anthropic's Model Context Protocol (MCP) for agent-to-tool access and Go

model-releasesarxiv-cs-ai
18 Aug 2026
Safety

Every Coin Has Two Sides: On the Dual Nature of Generalization in On-Policy Distillation of Large Language Models

DGX agent

arXiv:2608.16647v1 Announce Type: new Abstract: On-policy distillation (OPD) transfers teacher capabilities by supervising trajectories sampled from the student's own policy, yet its generalization be

safetyarxiv-cs-cl
18 Aug 2026
Model Releases

From Errors to Proofs: Minimal-Core-Guided Repair for Neuro-Symbolic Constraint Solving

DGX agent

arXiv:2608.14771v1 Announce Type: new Abstract: Making language models solve constraint problems reliably often means having them translate the problem into a formal specification and delegating the s

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Generated Context versus Governed State: Functional Conditions for Accountable Longitudinal Clinical Reasoning

DGX agent

arXiv:2608.14804v1 Announce Type: new Abstract: Large language models (LLMs) have become the dominant interface of clinical artificial intelligence, yet the interface they expose (text in, text out, o

model-releasesarxiv-cs-ai
18 Aug 2026
Local Ai

Is Ling 3 tiny underrated for its size?

DGX agent

I was checking out benchmarks of this model and apparantly better than Qwen3.5 9b reasoning across the bench on artificial analysis. I have used the 9b model for variety of stuff and it has been amazi

local-air-localllama
18 Aug 2026
Model Releases

MoE Router-Guided Clustering for Heterogeneous Federated Instruction Tuning

DGX agent

arXiv:2608.15311v1 Announce Type: new Abstract: Federated instruction fine-tuning enables Large Language Models (LLMs) to adapt to decentralized, privacy-sensitive data without requiring data sharing.

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

PixRestore: Unified Image Restoration via Pixel Diffusion Transformer

DGX agent

arXiv:2608.16793v1 Announce Type: new Abstract: Unified image restoration (UIR) aims to recover high-quality (HQ) content from low-quality (LQ) images with different degradations using a single model.

model-releasesarxiv-cs-cv
18 Aug 2026
Model Releases

Prompting is not enough: supervised baselines and leakage control for measuring shared decision-making with LLMs in pediatric encounters

DGX agent

arXiv:2608.14792v1 Announce Type: cross Abstract: Objectives: To determine whether zero-shot prompting of a large language model (LLM) is sufficient to detect shared decision-making (SDM) behaviors in

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Qwen 3.8 27b saved me $650+ in API costs this evening

DGX agent

I've been experimenting with Qwen3.8-27B using DeepSeek Harness. It's a monster at long-horizon tasks, and the results were pretty wild. DeepSeek Harness ran on my Windows PC and connected over LAN to

model-releasesr-localllama
18 Aug 2026
Model Releases

Qwen3.8 2.4T open weights made a Call of Duty clone

DGX agent

Qwen released the 2.4T Max weights and I was curious how well it can re-create COD in one prompt I ran the model on a rented B200 cluster and used roughly 1.1M output tokens over a 5 hour time span Re

model-releasesr-localllama
18 Aug 2026
Model Releases

Qwen3.8-27B on a 24GB M4 Pro Mac mini: benchmarks and the three settings that stop it drowning

DGX agent

When Qwen3.8-27B dropped on Thursday the obvious question came up for us : does a 27B dense model actually fit on the 24GB Mac Mini machines? Ran it properly over the weekend on an M4 Pro (24GB unifie

model-releasesr-ollama
18 Aug 2026
Model Releases

teams want to understand what their agents are doing but it’s hard to sift through all that trace data at scale but you can basically find a…

DGX agent

teams want to understand what their agents are doing but it’s hard to sift through all that trace data at scale but you can basically find any signal by fine-tuning a small, cheap model on data showin

model-releasesharrison-chase--x
18 Aug 2026
Model Releases

The Unwritten Benchmark: A New Challenge for Multimodal Machine Learning in Abstract Perceptual Reasoning

DGX agent

arXiv:2608.14558v1 Announce Type: new Abstract: Current multimodal models have demonstrated remarkable proficiency in recognizing static visual and auditory content. However, their capacity for abstra

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

The Working Set of a Coding Agent: Coherence Debt in Repository-Scale Tasks

DGX agent

arXiv:2608.16630v1 Announce Type: cross Abstract: Repository-scale coding requires an agent to keep tests, imports, configuration, and migration rules consistent within a bounded context window. We mo

model-releasesarxiv-cs-lg
18 Aug 2026
Safety

When State Becomes an Attack Surface: State-Semantic Injection in LLM-Driven Embodied Agents

DGX agent

arXiv:2608.16806v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated capabilities in in-context learning, task decomposition, step-by-step reasoning, and code generation, d

safetyarxiv-cs-ai
18 Aug 2026
Model Releases

When Stories Evolve: Benchmarking LLM Storytelling Across Agent Architectures in Open-Ended World Simulations

DGX agent

arXiv:2608.15654v1 Announce Type: cross Abstract: Large language models can write fluent stories, but open-ended storytelling requires more than local fluency. In evolving world simulations and AI-nat

model-releasesarxiv-cs-ai
18 Aug 2026
Research

A Data-Driven Algorithm for Model-Free Control Synthesis

DGX agent

arXiv:2602.13157v2 Announce Type: replace-cross Abstract: Presented is an algorithm to synthesize the optimal infinite-horizon LQR feedback controller for continuous-time systems. The algorithm does n

researcharxiv-cs-ro
17 Aug 2026
Model Releases

A Graph-Based Reinforcement Learning Framework for Structured Drift Diagnosis and Recovery in Autonomous LLM Agents

DGX agent

arXiv:2608.14109v1 Announce Type: new Abstract: Autonomous LLM agents are increasingly deployed in complex real-world workflows, yet they remain vulnerable to runtime behavioral drift, a silent deviat

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

Detecting Contaminated Code-Generation Prompt Batches via Influence Functions

DGX agent

arXiv:2608.14303v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for code generation, yet they remain vulnerable to prompts that elicit insecure implementations. Exis

model-releasesarxiv-cs-lg
17 Aug 2026
Model Releases

From single call to agents: five new Claude capabilities now available in Microsoft Foundry

DGX agent

Structured outputs, web search, web fetch, MCP connector, and tool search are now available for Claude models hosted on Azure in Microsoft Foundry, turning a model endpoint into a production agent pla

model-releasesmicrosoft-foundry
17 Aug 2026
Model Releases

How many tokens/second output are you getting with Qwen3.8-27B?

DGX agent

Trying to get a feel for where I stand. If you can list your relevant hardware and model used, that would be awesome. Here's mine: Model: Qwen3.8-27B-heretic-ara, Q5_K_M GGUF T/s: ~30-32 t/sec (I thin

model-releasesr-localllama
17 Aug 2026
Research

Non-Shattering at and Above the Dynamical Temperature in the Spherical Pure p-Spin Model

DGX agent

arXiv:2608.14369v1 Announce Type: cross Abstract: We consider the notion of shattering introduced by Ben Arous and Jagannath for spherical pure p-spin glasses with overlap q. For every pgeq 3 and 0sqr

researcharxiv-cs-lg
17 Aug 2026
Model Releases

NVIDIA Nemotron 3.5 Lightning now available in Amazon SageMaker JumpStart

DGX agent

NVIDIA Nemotron 3.5 Lightning, an open model built for high-volume agentic workloads, is now available in Amazon SageMaker JumpStart. This post shows how to deploy the 30B Mixture-of-Experts model (3B

model-releasesaws-ml-blog
17 Aug 2026
Research

S2Dialog: Multimodal Dialogue Retrieval with Semantic and Acoustic-Style Modeling

DGX agent

arXiv:2608.14029v1 Announce Type: new Abstract: Multimodal dialogue retrieval aims to retrieve dialogues from multimodal dialogue banks that are similar to a target dialogue in terms of both textual s

researcharxiv-cs-cl
17 Aug 2026
Safety

Some thoughts on Dario’s post: 1. Dario does not actually address Gavin Baker’s account of what he said – something he could easily deny if …

DGX agent

Some thoughts on Dario’s post: 1. Dario does not actually address Gavin Baker’s account of what he said – something he could easily deny if it were inaccurate. 2. Dario claims his critics live in a “b

safetyyann-lecun--x
17 Aug 2026
Model Releases

I am doing my best to stop back from this subject as it should be obvious what we are seeing here. However the arrogance that us commoners a…

DGX agent

I am doing my best to stop back from this subject as it should be obvious what we are seeing here. However the arrogance that us commoners are too dumb to catch his grift must be addressed. Receipts:

model-releasesyann-lecun--x
16 Aug 2026
Model Releases

Llama.cpp ROCm 7.2->7.14 upgrade, Radeon 780m iGPU benchmarks: ROCm vs Vulkan

DGX agent

With all the new models released recently one important upgrade went unnoticed: Llama.cpp bumped ROCm from 7.2 to 7.14. I was waiting for that because in 7.14 support for gfx1103 (Radeon 780m) was int

model-releasesr-localllama
16 Aug 2026
Model Releases

Newer commits removed the Qwen 35B

DGX agent

In this commits, the 35B model was removed. Looks like it's confirming the 35B model won't get released. I think they need to be made aware how big the 35 moe is widely used. Think need to make noise

model-releasesr-localllama
16 Aug 2026
Model Releases

Anthropic details Claude's text watermark: it only shows Claude was likely involved, is sparse in code and factual text, and disappears after a full rewrite (Anthropic)

DGX agent

Anthropic: Anthropic details Claude's text watermark: it only shows Claude was likely involved, is sparse in code and factual text, and disappears after a full rewrite — Future Claude models will gene

model-releasestechmeme
15 Aug 2026
Model Releases

Agent Behavioral Contracts II: Certifying Compositional Reliability Without Assuming Independence

DGX agent

arXiv:2608.12895v1 Announce Type: new Abstract: Compositional reliability bounds for multi-agent systems multiply component reliabilities, a step licensed by a conditional-independence assumption that

model-releasesarxiv-cs-ai
14 Aug 2026
Safety

Learning Under Treatment-Induced Label Indeterminacy with Expert Annotations of Counterfactual Outcomes: A Case Study in Neurological Prognostication

DGX agent

arXiv:2608.12477v1 Announce Type: new Abstract: Clinical prediction models are often developed as if the outcome of interest were cleanly observed for every patient. This assumption fails when treatme

safetyarxiv-cs-lg
14 Aug 2026
Model Releases

Qwen Live EP2 - Agent First: Multimodal Gets to Work - Streamed live 11 hours ago

DGX agent

Possibly best way to spend current hour before grabbing 27B model(Instead of opening duplicate repeated 27B posts here). I tried to grab summary(thought of including in this thread) of this video usin

model-releasesr-localllama
14 Aug 2026
Applications

RealMat: Realistic Materials with Diffusion and Reinforcement Learning

DGX agent

arXiv:2509.01134v2 Announce Type: replace-cross Abstract: Generative models for high-quality materials are particularly desirable to make 3D content authoring more accessible. However, the majority of

applicationsarxiv-cs-cv
14 Aug 2026
Model Releases

Rules or Character? Scaling Laws for AI Safety Design

DGX agent

arXiv:2608.13345v1 Announce Type: new Abstract: Artificial Intelligence (AI) safety systems combine character shaping (e.g., Reinforcement Learning from Human Feedback [RLHF], Constitutional AI), whic

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

Scaling Representation Diversity: Modulated Attention and Reconstructive Regularization for Visual Grounding

DGX agent

arXiv:2608.12748v1 Announce Type: new Abstract: Referring Expression Comprehension (REC) is commonly studied under dataset-specific fine-tuning, resulting in specialist models with limited cross-datas

model-releasesarxiv-cs-cv
14 Aug 2026
← Previous
1…343344345346347…1316
Next →