AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,092 results
Model Releases

Deciding When to Switch: E-Processes for Adaptive Minimax Training for Generative Adversarial Nets

DGX agent

arXiv:2608.10096v1 Announce Type: cross Abstract: Modern data science increasingly gives rise to hypothesis-testing problems that are not naturally formulated in terms of parameters within prespecifie

model-releasesarxiv-cs-lg
12 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

DeepSeek V4 Flash 0731 uncensored (jailbreak pt2)

DGX agent

Since lot's of people were sceptical or whatever, heres how to uncensor / jailbreak V4 flash and proof. No it is not lead on whatever, first prompt, first try, every time. Put this in System message:

model-releasesr-localllama
12 Aug 2026
Model Releases

DegradeQuery: Counterfactual Tuple Pretraining for Context-Aware PROTAC Degradation Prediction

DGX agent

arXiv:2608.10595v1 Announce Type: cross Abstract: Proteolysis-targeting chimeras (PROTACs) induce protein degradation by recruiting a target protein to an E3 ubiquitin ligase, making degradation a joi

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Derivative Computation in PINNs: Automatic Differentiation, Finite Differences and Beyond

DGX agent

arXiv:2608.11020v1 Announce Type: new Abstract: We systematically investigate finite-difference (FD) derivative computation in Physics-Informed Neural Networks (PINNs) as an alternative to automatic d

model-releasesarxiv-cs-lg
12 Aug 2026
Model Releases

Diffract: Spectral View of LLM Domain Adaptation

DGX agent

arXiv:2608.10850v1 Announce Type: new Abstract: We study continual pre-training (CPT) as a mechanism for adapting general-purpose large language models to specialized domains: mathematics, instruction

model-releasesarxiv-cs-lg
12 Aug 2026
Model Releases

DistilVDR: A Compact End-to-End Visual Document Retriever via Dual-Student Distillation

DGX agent

arXiv:2608.10636v1 Announce Type: cross Abstract: Visual document retrieval (VDR) is dominated by multi-billion-parameter models that are slow to index at full corpus scale and expensive to serve. Pri

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

Do LLM Recommenders Know When They're Hallucinating? Auditing Confidence Calibration in Catalog Faithfulness

DGX agent

arXiv:2608.10008v1 Announce Type: cross Abstract: LLM recommenders for top-K item suggestion regularly emit titles outside the target catalog. Prior audits measure this as a binary out-of-domain rate;

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

DreamOmni3: Scribble-based Editing and Generation

DGX agent

arXiv:2512.22525v2 Announce Type: replace Abstract: Recently unified generation and editing models have achieved remarkable success with their impressive performance. These models rely mainly on text

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

DriveVLA-M0: Failure-Aware Memory Augmentation for Autonomous Driving

DGX agent

arXiv:2608.10413v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for end-to-end autonomous driving by enabling unified reasoning across

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

DSAgentBench: Can Agents Automate End-to-End Data-Science Workflows in Real Computer Environments?

DGX agent

arXiv:2608.10366v1 Announce Type: new Abstract: Real-world data science involves long-horizon workflows that span data wrangling, exploration, modeling, visualization, and validation, and require coor

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

E^3mo-Bench: A Scalable Benchmark for Multimodal Evoked and Expressed Emotion Understanding via Bayesian Pairwise Alignment

DGX agent

arXiv:2608.10796v1 Announce Type: new Abstract: Understanding both expressed and evoked emotions is critical for multimodal large language models (MLLMs) to achieve comprehensive affect-aware interact

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

Edge Phoneme Recognition for Children's Speech through Age-Aware Training

DGX agent

arXiv:2608.10206v1 Announce Type: new Abstract: Detecting phonemes from children's speech has historically been difficult due to the scarcity of training data, and unique characteristics of children's

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Efficient Reinforcement Learning for Long-Horizon Tool-Use Agentic Tasks

DGX agent

arXiv:2608.10357v1 Announce Type: cross Abstract: Long-horizon tool-using agents must reason over user goals, domain policies, tool calls, simulator state, and delayed verifiable rewards. Reinforcemen

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

ENTLORE: A Graph-Grounded Benchmark for Latent Organizational Reasoning in Enterprise Question Answering

DGX agent

arXiv:2608.10679v1 Announce Type: cross Abstract: Enterprise question answering is framed as retrieving internal documents and generating grounded answers. Routine enterprise records, however, are wor

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Everyone else is talking about building ASI to like monopolize b2b saas and Elon is talking about building a kardashev II sentient sun

DGX agent

Everyone else is talking about building ASI to like monopolize b2b saas and Elon is talking about building a kardashev II sentient sun Media It’s always funny how people in SF twitter bubble will say

model-releaseselon-musk--x
12 Aug 2026
Model Releases

Evidence-Grounded Trustworthy Multimodal Reasoning and Evaluation Benchmark in Complex Urban Scenes

DGX agent

arXiv:2608.10954v1 Announce Type: cross Abstract: While Multimodal Large Language Models (MLLMs) demonstrate impressive performance in benign scenarios, their cognitive reliability deteriorates signif

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Exploration-Driven Personalized Federated Reinforcement Learning via Intrinsic Motivation

DGX agent

arXiv:2608.10499v1 Announce Type: cross Abstract: Personalized Federated Reinforcement Learning (PFRL) takes a decentralized approach to storing and accessing information based on past experiences whi

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Exploring Decoupled Spatio-Temporal Consistency Learning and Self-Prompting Evolution for Self-Supervised Tracking

DGX agent

arXiv:2507.21606v2 Announce Type: replace Abstract: The success of visual tracking has been largely driven by datasets with manual box annotations. However, these box annotations require tremendous hu

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

FADE: From Passive Verification to Active Discovery in Counterfactual Video Understanding

DGX agent

arXiv:2608.10764v1 Announce Type: new Abstract: Counterfactual video understanding evaluates whether models grasp physical and commonsense regularities. However, existing multiple-choice question (MCQ

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

FaithformBench: Benchmarking Faithfulness of Mathematical Chain-of-Thought Autoformalisation

DGX agent

arXiv:2608.10916v1 Announce Type: cross Abstract: Autoformalisation (AF) systems map natural language reasoning steps into formal statements in a proof assistant such as Lean. We consider how to asses

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Fast and Memory-Efficient Wavelet Convolutions via I/O-Aware Reformulation

DGX agent

arXiv:2608.10805v1 Announce Type: cross Abstract: Wavelet convolution (WTConv) has emerged as an increasingly popular drop-in replacement for standard convolutions, expanding a network's receptive fie

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Flex-pi: A Multi-Stream World-Action Model with Compute Flexibility

DGX agent

arXiv:2608.10860v1 Announce Type: cross Abstract: World-action models (WAMs) predict the future to act better, but nearly all of them predict only RGB latents, trained purely for pixel reconstruction,

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

FormStruct-Bench:A Hierarchical and Diagnostic Benchmark for Table-Form Document Structure Recognition

DGX agent

arXiv:2608.10396v1 Announce Type: new Abstract: Transforming table-form documents into machine-processable records requires recovering not only their visible content but also the multilevel structure

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

Four small architecture decisions can cost up to 47% of a model's long-context performance. New research from Ai2, Carnegie Mellon, and the …

DGX agent

Four small architecture decisions can cost up to 47% of a model's long-context performance. New research from Ai2, Carnegie Mellon, and the University of Washington isolates them. Normalization, GQA,

model-releasesdair-ai--x
12 Aug 2026
Model Releases

From #22 to #4 on Legal Research Bench! Solid progress for Qwen3.8-Max. Thanks for highlighting~✨

DGX agent

From #22 to #4 on Legal Research Bench! Solid progress for Qwen3.8-Max. Thanks for highlighting~✨ Qwen 3.8 Max nearly doubled its score on Legal Research Bench in under three months, climbing from #22

model-releasesqwen--x
12 Aug 2026
Model Releases

From Faulty Memories to Corrected Actions: Dependency-Guided Rollback Repair for Memory-Augmented Agents

DGX agent

arXiv:2608.10502v1 Announce Type: new Abstract: Persistent memory lets language-model agents reuse information across sessions, but it also makes errors durable: a poisoned, stale, or misattributed re

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

From Reasoning Depth to Reasoning Breadth: Evaluating Multi-Point Associative Reasoning in Large Language Models

DGX agent

arXiv:2608.10444v1 Announce Type: cross Abstract: Large language models (LLMs) have made substantial progress on reasoning tasks that require increasingly long and complex inferential chains. This pro

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

From Sync APIs to support for the GPT-5 model series and agentic workflows: What’s new in Azure Content Understanding – August 2026

DGX agent

Enterprise content is no longer just something people read. AI apps and agents are only as useful as the information they can understand, yet much of the world’s enterprise knowledge is locked in docu

model-releasesmicrosoft-foundry
12 Aug 2026
Model Releases

Gemma 4 QAT handles KV cache quantization MUCH better, KLD benchmarks show

DGX agent

Link to the article: KV Cache Quantization on Gemma 4 31B: Non-QAT vs QAT KLD benchmarks with BeeLlama.cpp v0.4.3, fork of llama.cpp with more KV cache quantization options, comparing Gemma Q4_0 non-Q

model-releasesr-localllama
12 Aug 2026
Model Releases

GeoForge: Non-Parametric Self-Evolving Agents for Earth-Observation Reasoning

DGX agent

arXiv:2608.10494v1 Announce Type: new Abstract: Earth observation (EO) agents construct scientifically valid tool workflows and ground their conclusions in current geospatial evidence. This is challen

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

GeoSeg-OV: Bridging Geospatial Gaps with Structural Guidance for Open-Vocabulary Remote Sensing Segmentation

DGX agent

arXiv:2608.10426v1 Announce Type: new Abstract: Open-vocabulary remote sensing segmentation has recently emerged as a promising paradigm that enables pixel-level recognition of arbitrary categories sp

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes

DGX agent

arXiv:2608.10886v1 Announce Type: new Abstract: Robots operating in human environments need memories that capture not only what objects exist and where, but also how people use them over time and how

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

give 4.6 a try and let us know how it goes. your feedback is a big part of why the model gets better with each iteration.

DGX agent

give 4.6 a try and let us know how it goes. your feedback is a big part of why the model gets better with each iteration. SpaceXAI's Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index, j

model-releaseselon-musk--x
12 Aug 2026
Model Releases

GLAM: Efficient Continual Learning at Scale via Grouped LoRA Adapter Merging

DGX agent

arXiv:2509.13211v4 Announce Type: replace Abstract: The ability to learn continuously over time remains a major challenge for modern machine learning systems, even in the era of Foundation Models. Whi

model-releasesarxiv-cs-lg
12 Aug 2026
Model Releases

Google DeepMind launches SL2T, a multilingual sign-language-to-text model debuting on the Pixel 11 in Gboard and Live Transcribe, first with ASL and English (Mike Wheatley/SiliconANGLE)

DGX agent

Mike Wheatley / SiliconANGLE: Google DeepMind launches SL2T, a multilingual sign-language-to-text model debuting on the Pixel 11 in Gboard and Live Transcribe, first with ASL and English — Google Deep

model-releasestechmeme
12 Aug 2026
Model Releases

Google unveils the $399 Pixel Watch 5 with a satin pyrite case finish, offline Gemini, proactive AI suggestions, better GPS maps, and insulin resistance trends (Victoria Song/The Verge)

DGX agent

Victoria Song / The Verge: Google unveils the 399 Pixel Watch 5 with a satin pyrite case finish, offline Gemini, proactive AI suggestions, better GPS maps, and insulin resistance trends — The 399 Goog

model-releasestechmeme
12 Aug 2026
Model Releases

Google unveils the 899+ Pixel 11, 1,099+ 11 Pro, and $1,299+ 11 Pro XL, with a Tensor G6, new Gemini features, Magic Capture to pick the best frames, and more (Ivan Mehta/TechCrunch)

DGX agent

Ivan Mehta / TechCrunch: Google unveils the 899+ Pixel 11, 1,099+ 11 Pro, and $1,299+ 11 Pro XL, with a Tensor G6, new Gemini features, Magic Capture to pick the best frames, and more — For the last f

model-releasestechmeme
12 Aug 2026
Model Releases

Grok 4.6 is an excellent model. I’ve been using it heavily for the past couple of weeks and it handles everything from simple coding & code …

DGX agent

Grok 4.6 is an excellent model. I’ve been using it heavily for the past couple of weeks and it handles everything from simple coding & code review all the way to designing and debugging complex system

model-releaseselon-musk--x
12 Aug 2026
Model Releases

Grok 4.6 is now one of the top models in the world for agentic workflows It ranks #1 on the Artificial Analysis Agentic Index, tied with Cla…

DGX agent

Grok 4.6 is now one of the top models in the world for agentic workflows It ranks #1 on the Artificial Analysis Agentic Index, tied with Claude Opus 5 Max Agentic AI is about more than answering quest

model-releaseselon-musk--x
12 Aug 2026
Model Releases

Grok 4.6 is objectively #1 when considering intelligence, speed & cost

DGX agent

Grok 4.6 is objectively #1 when considering intelligence, speed & cost SpaceXAI's Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index, joining the frontier in line with GPT-5.6 Sol, with

model-releaseselon-musk--x
12 Aug 2026
Model Releases

Grok 4.6 reaches 1753 ELO

DGX agent

Grok 4.6 topped the GDPVal-AA benchmark with an Elo score of 1,753. It surpassed competitors Fable 5 Max (1,741 Elo), GPT‑5.6 Sol Max (1,728 Elo) and Grok 4.5 High (1,526 Elo). Elon Musk publicly ackn

model-releaseselon-musk--x
12 Aug 2026
Model Releases

Grok Bot

DGX agent

Grok Bot Here's my Grok Bot team: - Webby: Web designer - Shotry: Short-form content creator - Writey: Article/Newsletter writer - Claude Code: Grok agent that specializes in CC - Codex: Same as the a

model-releaseselon-musk--x
12 Aug 2026
Model Releases

Hidden Reasoning from Claude and GPT are Decoded, and it is interesting

DGX agent

Yesteday a paper showed a gap that allows to see 100% of the reasoning tokens form ALL Claude and GPT models Stealing Reasoning Traces from Proprietary LLM APIs. check it out, they have published lots

model-releasesr-localllama
12 Aug 2026
Model Releases

HNDiff: Haze-Noise Diffusion for Image Dehazing

DGX agent

arXiv:2608.10995v1 Announce Type: new Abstract: Existing diffusion-based methods have recently made significant progress in image dehazing. However, they typically neglect the physics of haze formatio

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

HoosierHelp: Benchmarking LLM Agents for Social Service Navigation

DGX agent

arXiv:2608.09946v1 Announce Type: cross Abstract: Social service navigation requires connecting help-seeking individuals to resources that satisfy their needs and specific constraints. Although LLM ag

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

How OneAdvanced deployed over 50 AI agents on UK-sovereign AWS

DGX agent

Learn how OneAdvanced, a UK enterprise software provider, built a UK-sovereign AI platform by self-hosting Llama 4 Maverick and Llama Guard 4 on Amazon SageMaker AI, with a RAG pipeline on pgvector an

model-releasesaws-ml-blog
12 Aug 2026
Model Releases

How Robust Are LLMs to Vietnamese Dialects?

DGX agent

arXiv:2608.10414v1 Announce Type: new Abstract: Large Language Models (LLMs) are typically evaluated on standard written Vietnamese, yet everyday communication frequently involves regional dialects th

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

DGX agent

arXiv:2506.03922v4 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated significant potential to advance a broad range of domains. However, current benchma

model-releasesarxiv-cs-ai
12 Aug 2026
← Previous
12345…461
Next →