AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,292 results
Model Releases

DelusionEval: Measuring Delusion-Linked Behaviors in AI Chatbots

DGX agent

arXiv:2608.05004v1 Announce Type: new Abstract: Mental health professionals have raised concerns about risks of psychological harm from interaction with large language models (LLMs), including 'delusi

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Diagnosing Tool-Selection Reasoning in LLM Agents with Canary Tools

DGX agent

arXiv:2608.04719v1 Announce Type: new Abstract: Agent evaluations tell us that a model picked the wrong tool, but rarely why. We introduce canary tools: diagnostic probe tools planted in an agent's Mo

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Differential 6-DOF Pose Estimation with Provable First-Order Immunity to Camera Calibration Errors

DGX agent

arXiv:2608.04673v1 Announce Type: new Abstract: Accurate six-degree-of-freedom (6-DOF) motion estimation is essential for robotic manipulation, autonomous systems, and structural displacement monitori

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Digital sovereignty in the age of AI: You don’t have to choose between control and innovation

DGX agent

For enterprises and governments with strict compliance and sovereignty requirements, keeping sensitive data on-premises often means missing out on the latest AI. These organizations are managing three

model-releasesgoogle-cloud-ai
6 Aug 2026
Model Releases

Diverse and Plausible Algorithmic Recourse via Tractable Recourse Distributions

DGX agent

arXiv:2608.04677v1 Announce Type: new Abstract: Algorithmic recourse seeks to help individuals reverse unfavorable automated decisions by recommending actionable changes that achieve a desired outcome

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

DreamWAM: Beyond RGB Future Prediction for World Action Models

DGX agent

arXiv:2608.04996v1 Announce Type: new Abstract: World Action Models (WAMs) learn action-relevant representations by predicting how the observed world will evolve. Most existing WAMs define this future

model-releasesarxiv-cs-ro
6 Aug 2026
Model Releases

Dual 3090 setup: 400 pp t/s to 1600 pp t/s on Qwen 3.6 27B... with slightly lower tps.

DGX agent

First of all, my setup: Ryzen 9 5950x DDR4 3200Mhz 64gb (2x32) Dual 3090s, no NVLINK Runtime: llama.cpp Nvidia Drivers 610 Windows 11 25H2 Qwen 3.6 27B Q8 I've been using llama-server with --split-mod

model-releasesr-localllama
6 Aug 2026
Model Releases

Dynamic Jailbreaking Attack

DGX agent

arXiv:2510.02422v4 Announce Type: replace-cross Abstract: Existing gradient-based jailbreak attacks typically optimize a fixed-length adversarial suffix toward a predefined target response with a stat

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Easy to Complete, Hard to Choose: Investigating LLM Performance on the ProverbIT Benchmark

DGX agent

arXiv:2608.04670v1 Announce Type: cross Abstract: Large Language Models (LLMs) have transformed computational linguistics and achieved remarkable performance across numerous natural language processin

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Echo Flow Networks

DGX agent

arXiv:2509.24122v3 Announce Type: replace Abstract: At the heart of time-series forecasting (TSF) lies a fundamental challenge: how can models efficiently and effectively capture long-range temporal d

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

EDATracer: An Agentic Framework for Large-Scale EDA Artifact Analysis

DGX agent

arXiv:2608.04032v1 Announce Type: cross Abstract: Modern chip design relies on electronic design automation (EDA) tools that generate large, heterogeneous artifacts, including source files, scripts, l

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Efficient Online Lexicographic Generalized Low-Rank Matrix Bandits

DGX agent

arXiv:2608.04324v1 Announce Type: cross Abstract: This paper studies generalized low-rank matrix bandits with multiple prioritized objectives. At each round, the learner selects a matrix-valued arm an

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

EgoAfford: Task-Oriented Affordance Grounding via Egocentric Referring Segmentation

DGX agent

arXiv:2608.04533v1 Announce Type: new Abstract: Part-level affordance grounding has advanced the localization of functional object regions associated with elemental actions. Extending this capability

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Eliciting Intrinsic Hallucinations in LLMs via Semantically Equivalent Adversarial Attacks

DGX agent

arXiv:2608.04286v1 Announce Type: new Abstract: Large language models (LLMs) are often used in conjunction with external knowledge sources to improve their factual accuracy and decrease hallucinations

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Energy- and Memory-Efficient PEFT Methods for Personalized On-Device SLMs on Consumer GPUs

DGX agent

arXiv:2608.04488v1 Announce Type: new Abstract: Despite rapid advances in large language models (LLMs), deploying and personalizing them on resource-constrained devices remains impractical due to high

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Energy-Tweedie: Score meets Score, Energy meets Energy

DGX agent

arXiv:2512.23818v2 Announce Type: replace-cross Abstract: Denoising and score estimation are classically linked through Tweedie's formula, which relates the posterior mean under Gaussian noise to the

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Enforcing data residency with single-Region Claude Code on Amazon Bedrock

DGX agent

A regulated customer needed all Claude Code inference processed in a single AWS Region (London), not just in-geography. This post shows two ways to pin Claude Code on Amazon Bedrock to one Region: an

model-releasesaws-ml-blog
6 Aug 2026
Model Releases

Enhancing Trustworthy Clinical Diagnosis Decision-Making in Large Language Models via Etiology-Aware Attention Supervision

DGX agent

arXiv:2508.00285v2 Announce Type: replace Abstract: Objective: Large Language Models (LLMs) have demonstrated strong capabilities in medical text understanding and generation. However, their trustwort

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

EuroExec: Frontier Language Models Fall Short of Expert Judgment on European Executive Decision Tasks

DGX agent

arXiv:2608.04549v1 Announce Type: cross Abstract: Frontier LLMs are increasingly put to use on open-ended complex questions, different in nature from the ones they are typically evaluated on. We dedic

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

ExeCRE: Execution-Consistency Guided Reliability Estimation for Self-Correcting Code Generation

DGX agent

arXiv:2608.04439v1 Announce Type: cross Abstract: Large language models (LLMs) have made notable progress in code generation, but they still struggle on challenging tasks that require sophisticated al

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Faster-WAM: Efficient Inference-Time Future Conditioning for Robust World Action Models

DGX agent

arXiv:2608.04404v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot manipulation by learning how the environment evolves beyond the current observation. However, existing approach

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Filing: DeepSeek has invested ~$20.8M in Unitree Robotics' Shanghai IPO and agreed to jointly develop AI models for humanoid machines (Eduardo Baptista/Reuters)

DGX agent

Eduardo Baptista / Reuters: Filing: DeepSeek has invested ~20.8M in Unitree Robotics' Shanghai IPO and agreed to jointly develop AI models for humanoid machines — Chinese artificial intelligence start

model-releasestechmeme
6 Aug 2026
Model Releases

Final optimization: from ~10 tok/s to ~15 tok/s on DeepSeek-V4-Flash-0731 at 128K ctx - 1 RTX 3090

DGX agent

J'ai consacré beaucoup de temps à l'optimisation de DeepSeek-V4-Flash-0731 GGUF sur une seule RTX 3090. Mon exigence absolue pour chaque configuration était la suivante : Le modèle doit rester utilisa

model-releasesr-localllama
6 Aug 2026
Model Releases

FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM Agents

DGX agent

arXiv:2608.04095v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used as personalized assistants in high-stakes domains such as financial advising, yet it remains unc

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables

DGX agent

arXiv:2608.04077v1 Announce Type: new Abstract: Evaluating financial AI agents requires criteria aligned with real professional work. Existing rubric methods typically derive criteria from task prompt

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation

DGX agent

arXiv:2608.04374v1 Announce Type: cross Abstract: Large language models can produce fluent financial analysis, but fluency alone does not establish whether a report is suitable for institutional deliv

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

FinRpt: Dataset, Evaluation System and LLM-based Multi-agent Framework for Equity Research Report Generation

DGX agent

arXiv:2511.07322v3 Announce Type: replace-cross Abstract: While LLMs have shown great success in financial tasks like stock prediction and question answering, their application in fully automating Equ

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Foreseeing the Invisible: Amodal Reconstruction of Leaf Fossil Images

DGX agent

arXiv:2608.04423v1 Announce Type: new Abstract: Fossil leaves are rarely preserved whole -- sedimentary rock hides, breaks, and erodes the lamina, yet paleobotany depends on the complete shape and out

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Formal Analysis and Supply Chain Security for Agentic AI Skills

DGX agent

arXiv:2603.00195v2 Announce Type: replace-cross Abstract: 32 pages, 5 theorems with full proofs, 68 references, open-source tool: https://github.com/qualixar/skillfortify. v2: corrects the bibliograph

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

From Financial Sentiment Classification to Return Predictability: A QLoRA Benchmark of Large Language Models

DGX agent

arXiv:2608.04200v1 Announce Type: cross Abstract: Financial sentiment classifiers are commonly evaluated against human labels, but strong linguistic performance does not necessarily imply economically

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

From Score Matrices to Football-Aware Match-State Simulation: An Auditable LLM Harness for Exact-Score Reranking

DGX agent

arXiv:2608.05030v1 Announce Type: new Abstract: Football score forecasting combines a strong statistical core with a difficult contextual edge. Dynamic Poisson-family models estimate team strength, ex

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

From Transparent Labware Segmentation to Collision Avoidance: A Real-Time Edge-Aware Perception Pipeline

DGX agent

arXiv:2608.04769v1 Announce Type: new Abstract: This paper presents an edge-aware instance segmentation framework that enables real-time robotic collision avoidance with transparent laboratory glasswa

model-releasesarxiv-cs-ro
6 Aug 2026
Model Releases

FUSEP: A Multi-Center Benchmark for Diverse Tasks in Early Pregnancy Fetal Ultrasound Screening

DGX agent

arXiv:2608.04766v1 Announce Type: cross Abstract: A large number of infants with congenital anomalies are born each year globally, especially in areas with underdeveloped medical resources. Currently,

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Galaxy Phase-Space and Field-Level Cosmology: The Strength of Semi-Analytic Models

DGX agent

arXiv:2512.10222v2 Announce Type: replace-cross Abstract: Semi-analytic models are a widely used approach to simulate galaxy properties within a cosmological framework, relying on simplified yet physi

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

GEB-Bench: Abstract Structures Told in Many Voices

DGX agent

arXiv:2608.04111v1 Announce Type: cross Abstract: Can a model look at a river delta and a lightning bolt and see that they share a structure? We introduce GEB-Bench, a benchmark whose unit is an abstr

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

General Availability of Pinecone Nexus Proves Knowledge Drives Real Outcomes for Agentic AI

DGX agent

Pinecone announced the general availability of Pinecone Nexus, a knowledge engine that converts an enterprise’s proprietary data into governed, agent‑ready knowledge delivered through a single query c

model-releasespinecone
6 Aug 2026
Model Releases

Geometry-Informed Parameter-Efficient Fine-Tuning of Pre-trained Molecular GNNs for Blood-Brain Barrier Permeability Prediction

DGX agent

arXiv:2608.04257v1 Announce Type: new Abstract: Blood-brain barrier permeability (BBBP) prediction is a critical screening task in central nervous system drug discovery, where candidate molecules must

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

GLM/Qwen Appreciation Post

DGX agent

https://preview.redd.it/o6ik6qboeohh1.png?width=1134&format=png&auto=webp&s=4016f26c50c1d93bd3d0c7e880e9b55a2d75310f I have been running Qwen3.6 27b for a little while (mostly coding tasks) and recent

model-releasesr-localllama
6 Aug 2026
Model Releases

Google is expanding its Gemini-powered Ask Maps feature in English to 130+ countries and adds agentic food ordering, hotel booking, and attraction searching (Jada Jones/ZDNET)

DGX agent

Jada Jones / ZDNET: Google is expanding its Gemini-powered Ask Maps feature in English to 130+ countries and adds agentic food ordering, hotel booking, and attraction searching — ZDNET's key takeaways

model-releasestechmeme
6 Aug 2026
Model Releases

@GoogleDeepMind Humanoid legs or wheeled rovers? Should robots be cracking eggs? Watch as the @GoogleDeepMind team behind Gemini Robotics 2 …

DGX agent

On July 30 Google AI published a video announcing **Gemini Robotics 2**, an intelligence layer developed by DeepMind that aims to bring autonomous robots closer to everyday human environments. The pos

model-releasesgoogle-ai--x
6 Aug 2026
Model Releases

Gradient Immunity: Null-Space Resistance to Malicious Fine-Tuning

DGX agent

arXiv:2608.05045v1 Announce Type: cross Abstract: Released aligned large language models remain vulnerable to malicious downstream finetuning. Existing defenses are largely designed for the fine-tunin

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

GRALS: GCN-Guided Redundancy-Aware Local Search for Minimum Vertex Cover

DGX agent

arXiv:2503.06396v2 Announce Type: replace Abstract: The minimum vertex cover (MVC) problem seeks to identify the smallest set of vertices that cover all edges in an undirected graph. As a fundamental

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

GUARD: Grounding Uncertainty and Ablation-Based Risk Detection for Diffusion-Based VLAs

DGX agent

arXiv:2608.04510v1 Announce Type: cross Abstract: Diffusion-based vision-language-action (VLA) policies can generate plausible actions even when their predictions are weakly grounded in the visual and

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Hallucinations on the Board: Tool-Augmented Evaluation of LLM Chess Commentary

DGX agent

arXiv:2608.04240v1 Announce Type: cross Abstract: Superhuman game engines in domains like chess have made expert-level evaluations easily accessible, yet they communicate what is true without the natu

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

HelloWorld: Enabling Socially Interactive Characters in Video World Models

DGX agent

arXiv:2608.05070v1 Announce Type: new Abstract: Despite the remarkable recent progress of video world models, social interaction between users and the characters within these worlds remains unsupporte

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation

DGX agent

arXiv:2608.04378v1 Announce Type: cross Abstract: Collaborative music agents need internal representations rich enough to support both understanding and generation, yet flexible enough for a workflow

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

How come artificialanalysis.ai ranks Gemma4 above Qwen3.6 27b in SciCode

DGX agent

Just came across this coding benchmark: SciCode Artificialanalysis.ai reports a ranking which contradicts the feeling we've towards those models in real life coding. Is Gemma 4 really that good, or a

model-releasesr-localllama
6 Aug 2026
Model Releases

How is Deepseek v4 flash 0731 running on Ollama cloud?

DGX agent

I cancelled my pro plan ealier because I wanted to use new Deepseek v4 flash 0731 which was available on Openrouter through API only (not yet on ollama cloud at the time). The old Deepseek v4 flash/pr

model-releasesr-ollama
6 Aug 2026
← Previous
1…3132333435…465
Next →