AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,292 results
5 Aug 2026

Trajectory inference via Acceleration Matching

Model ReleasesDGX agent

arXiv:2608.03916v1 Announce Type: new Abstract: Trajectory inference is a fundamental problem in many scientific domains: given a collection of unpaired snapshots of observations at discrete time poin

Try npm i -g cline on Cline!👀

Model ReleasesDGX agent

Try npm i -g cline on Cline!👀 Qwen3.8-Max is now available in ClinePass, a subscription for ~5x discounted access. Use it on Cline CLI w/ $4.99 special promo: npm i -g cline This is currently the most

TumorBoard: Evidence-Grounded Multi-Agent Decision Support for Longitudinal Neuro-Oncology

Model ReleasesDGX agent

arXiv:2608.03190v1 Announce Type: new Abstract: Neuro-oncology decisions require coordinated interpretation of serial MRI, pathology, molecular markers, treatment history, performance status, and evol


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Unequal Verdicts: Investigating Gender Bias in LLM-Based Fake News Detection

Model ReleasesDGX agent

arXiv:2608.03627v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used for automated fact-checking, yet their susceptibility to gender bias in this context remains underexp

UniNav: A Unified World-Action Diffusion Model for Visual Navigation

Model ReleasesDGX agent

arXiv:2608.03244v1 Announce Type: new Abstract: Image-goal visual navigation is a fundamental capability for embodied agents. Existing navigation policies efficiently predict waypoint trajectories but

UniWorld-Design: From Pixel Generation to Layer-Native Design

Model ReleasesDGX agent

arXiv:2608.03971v1 Announce Type: new Abstract: We introduce UniWorld-Design, a framework that redefines image generation from flat pixel synthesis to structured visual composition, with semantic RGBA

Unlocking the future of shared storage: Filestore on Colossus

Model ReleasesDGX agent

Today, enterprise storage must be as agile, elastic, and responsive as the workloads it supports. Filestore, Google Cloud’s first-party, secure, scalable NFS file service, can service a wide-range of

UrbanAgent: A Tool-Augmented Agent for Cross-System Urban Tasks

Model ReleasesDGX agent

arXiv:2608.03018v1 Announce Type: new Abstract: Modern cities rely on an increasing number of digital services to operate, but residents' daily needs are still difficult to meet. Services are fragment

Utilize a nvidia gpu and amd gpu together for 2 different ai models?

Model ReleasesDGX agent

We run a local model instance in our company that the dev we hired built for us. We're a trade business and we want to further use our on hand hardware for it. The specs given we have is a 5090 gpu wi

Vectra AI launches Vectra AI Pro to feed AI agents better attack signals

Model ReleasesDGX agent

Threat detection company Vectra AI Inc. today launched Vectra AI Pro, a product built to give the artificial intelligence agents now working inside security operations centers a more reliable read on

Verifier-Guided Model Discovery for Physical Dynamical Systems with Pretrained Symbolic Transformers

Model ReleasesDGX agent

arXiv:2608.02662v1 Announce Type: cross Abstract: Reliable forecasting of nonlinear physical systems underpins scientific discovery and engineering decision-making. Yet high-fidelity simulations are p

VeriTrace: Human-Like Temporal Exploration Completes Agentic Action Space

Model ReleasesDGX agent

arXiv:2608.02878v1 Announce Type: new Abstract: Large language models have shown promise for automated Verilog RTL generation, yet state-of-the-art multi-agent systems plateau at ~95% accuracy on stan

VIBE: A VAD-Informed Benchmark for Entity-Centered Affective Profiling of Large Language Model Outputs

Model ReleasesDGX agent

arXiv:2608.03810v1 Announce Type: cross Abstract: Large language models routinely describe socially salient targets, including political figures, countries, religions, organizations, historical events

VIBE: Vector Index Benchmark for Embeddings

Model ReleasesDGX agent

arXiv:2505.17810v2 Announce Type: replace Abstract: Approximate nearest neighbor (ANN) search is a performance-critical component of many machine learning pipelines, and rigorous benchmarking is essen

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent

Model ReleasesDGX agent

arXiv:2608.03979v1 Announce Type: cross Abstract: We introduce Video-DeepResearch (Video-DR), extending multimodal agents from static images to continuous video streams, a setting that demands dense s

VIVID: A Culturally Grounded Benchmark Exposing the Figurative Language Gap in Vietnamese NLP

Model ReleasesDGX agent

arXiv:2608.03095v1 Announce Type: new Abstract: We present VIVID (Vietnamese Idioms for Validation and Interpretation Depth), the first systematic benchmark for evaluating culturally grounded figurati

Vulnerabilities, Secrets and Misconfiguration in the Highest-Exposure Docker Hub Images

Model ReleasesDGX agent

arXiv:2608.02669v1 Announce Type: cross Abstract: Docker Hub is the registry underneath most container deployments, and a flaw in a widely reused base image is inherited by every image built on it. Pr

Watch a local Ollama's qwen3:8b turn one English question into a 9-node investigation graph - planned, admitted by a deterministic gate, and run live in the browser (open source, MIT)

Model ReleasesDGX agent

The video is one real run, not a mock-up: grapharc go 'why did checkout latency spike at 09:14 UTC?' --model ollama/qwen3:8b A local 8B model proposes the graph → triage fanning out into four parallel

WeClawArena: An Auditable Sandbox and Benchmark for Cross-User Agents Collaboration and Security in Human-Centered Agent Networks

Model ReleasesDGX agent

arXiv:2608.03499v1 Announce Type: new Abstract: Recent advances in persistent personal-agent frameworks are making human-centered agent networks realistic deployment targets: each user can be served b

When Agents Learn to Be You: Benchmarking Privacy Leakage, Impersonation Risk, and Defenses in Persona Skills

Model ReleasesDGX agent

arXiv:2608.03700v1 Announce Type: cross Abstract: Persona skills distill personal interaction histories into portable and executable artifacts for downstream agents. While enabling flexible personaliz

When Attention Goes Blind: Numerical Failure in ALiBi Positional Encodings

Model ReleasesDGX agent

arXiv:2608.03994v1 Announce Type: new Abstract: We identify a previously overlooked failure mode of ALiBi positional encoding: its linear bias scaling underflows floating-point precision, which zeroes

When Classes Evolve: A Benchmark and Framework for Stage-Aware Class-Incremental Learning

Model ReleasesDGX agent

arXiv:2602.00573v2 Announce Type: replace-cross Abstract: Class-Incremental Learning (CIL) aims to sequentially learn new classes while mitigating catastrophic forgetting of previously learned knowled

When Correct Solutions Repeat: Rarity-Aware Credit Redistribution for GRPO

Model ReleasesDGX agent

arXiv:2608.03467v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) com- monly optimizes each correct completion as an independent learning signal. In GRPO, this comp

When Do Fewer Visual Tokens Accelerate Multimodal Inference? A Break-Even Study Across Decision Locations and Hardware

Model ReleasesDGX agent

arXiv:2608.03649v1 Announce Type: new Abstract: Fewer visual tokens do not guarantee lower end-to-end latency. We evaluate break-even with a reproducible protocol that accounts for decision overhead,

When Many Answers Are Valid, Voting Fails: Symbolic Verification for Best-of-K Causal Reasoning in LLMs

Model ReleasesDGX agent

arXiv:2608.03506v1 Announce Type: new Abstract: Self-consistency assumes the most frequent answer among sampled reasoning traces is the most reliable, but this can fail in causal reasoning: samples of

When Memory Becomes Authority: Benchmarking Authority Collapse at the Memory Consolidation Boundary

Model ReleasesDGX agent

arXiv:2608.01679v2 Announce Type: replace Abstract: Persistent memory allows (self-evolving) LLM agents to adapt across tasks by consolidating heterogeneous interaction histories into reusable facts,

When Outputs Disperse, Does Epistemic Revision Follow? A Black-Box Coupling Diagnostic for Machine Collectives

Model ReleasesDGX agent

arXiv:2608.03722v1 Announce Type: new Abstract: Collective intelligence research treats disagreement as evidence of epistemic diversity: if agents express different views, the group should retain capa

When Refusal Looks Safe: The Refusal-Cue Shortcut in Safety Guard Models

Model ReleasesDGX agent

arXiv:2608.03201v1 Announce Type: new Abstract: Safety guards are widely used to filter harmful content and are typically trained via supervised fine-tuning on labeled prompt-response pairs. We audit

Where Did It Go Wrong? Process-Level Evaluation of Web Agents with Semantic State Tracking

Model ReleasesDGX agent

arXiv:2606.15673v2 Announce Type: replace Abstract: Web agents act through long interaction sequences, yet existing benchmarks evaluate only terminal success, discarding all process information and of

Wiring Beats Blending: What Transfers Between Transformer Sizes -- and What Doesn't

Model ReleasesDGX agent

arXiv:2608.02829v1 Announce Type: new Abstract: Model families train every size from scratch. Can a pretrained large model be converted into a smaller sibling? We characterize the 1.4B->410M conversio

WorldCup Arena: Prospective, Leakage-Free Evaluation of Frontier LLMs on a Live Tournament

Model ReleasesDGX agent

arXiv:2608.04008v1 Announce Type: new Abstract: Benchmarks that measure the forecasting ability of large language models are almost always retrospective: the event has happened, the answer is somewher

Xiaomi-Robotics-1: New robotics model released

Model ReleasesDGX agent

Xiaomi-Robotics-1 is a robot foundation model trained on over 100K hours of real-world manipulation trajectories. It is a Vision-Language-Action (VLA) model engineered for out-of-the-box mobile manipu

You have Dwarkesh predicting Anthropic is going to make well over 100B in revenue this year. and then you have this chart.

Model ReleasesDGX agent

Gary Marcus noted that investor Dwarkesh predicts Anthropic will generate more than $100 billion in revenue this year, according to a source. A chart accompanying the claim omitted DeepSeek’s pricing

Your Agentic LLMs Secretly Encode Latent Signals of Indirect Prompt-Injection Exposure

Model ReleasesDGX agent

arXiv:2608.02657v1 Announce Type: cross Abstract: Agentic LLMs are vulnerable to indirect prompt injection (IPI) attacks, e.g., malicious side-tasks hidden in external tool results. While many efforts

4 Aug 2026

A 2.6B model with tool calling and 128K context now runs at 30 tok/s on a phone

Model ReleasesDGX agent

Liquid AI released LFM2.5-2.6B today, and this might be more relevant to local AI than another massive model most people cannot run. The model is only 2.69B parameters, has 128K context, supports tool

A Benchmark Dataset for MLLM-Generated Image Detection: GPT Image2 & Nano Banana2

Model ReleasesDGX agent

arXiv:2608.01258v1 Announce Type: new Abstract: The realism of images generated by multimodal large language models (MLLMs), such as GPT Image2 and Nano Banana2, has improved rapidly in recent years.

A Comparative Analysis of MLP and Kolmogorov-Arnold Networks (KAN) for Faster-than-Nyquist (FTN) Signaling Detection

Model ReleasesDGX agent

arXiv:2608.02062v1 Announce Type: cross Abstract: Faster-than-Nyquist signaling improves spectral ef- ficiency by deliberately introducing inter-symbol interference. Classical sequence detectors such

A Few Neurons Reveal When LLMs Misuse Tools: Sparse Detection and Selective Steering for Reliable Tool Use

Model ReleasesDGX agent

arXiv:2608.00218v1 Announce Type: new Abstract: Agentic LLMs exhibit three consequential tool-use failures: invalid arguments (validity), unnecessary calls (over-calling), and omitted calls when tools

A General-Purpose VLM Can Teach an Astronomy Foundation Model to Better Recognize Galaxy Morphology

Model ReleasesDGX agent

arXiv:2608.02300v1 Announce Type: new Abstract: Existing astronomy foundation models provide strong galaxy representations, but adapting them to new survey conditions and survey-specific morphology re

A Large-Scale Multi-Dimensional Empirical Study of LLMs for Conversation Summarization

Model ReleasesDGX agent

arXiv:2606.15974v2 Announce Type: replace Abstract: Despite the significant advancement of LLMs in conversation summarization, their evaluation remains limited by insufficient scenarios, input lengths

A llama.cpp PR caches “hot” MoE experts on the GPU — 33 → 56 tok/s reported with 8GB VRAM

Model ReleasesDGX agent

A new llama.cpp PR (#26563) adds a heatmap that tracks which MoE experts are used most often. Instead of keeping every expert on the GPU or offloading all of them, it caches the frequently selected ex

A Simple Approximation to the Distribution of the Ridge Regression Estimator

Model ReleasesDGX agent

arXiv:2608.02539v1 Announce Type: cross Abstract: We present a simple Gaussian approximation to the finite-sample distribution of the classical ridge regression estimator. Our approximation captures t

A Triple-Robustness Analysis of Retrieval-Augmented Generation for Multi-Hop Requirements Traceability

Model ReleasesDGX agent

arXiv:2608.00705v1 Announce Type: cross Abstract: Reported verdicts on GraphRAG versus vector RAG disagree, and the evidence is typically tied to a single corpus, embedder, and judge -- and, we show,

Abduction Without a Body? Representational Grounding and the Abduction Loop for Scientific Hypothesis Generation

Model ReleasesDGX agent

arXiv:2608.02505v1 Announce Type: cross Abstract: Can scientific abduction occur without continuous sensorimotor embodiment? Recent arguments in AI and philosophy of science hold that genuine hypothes

Accuracy Does Not Guarantee Human-Likeness: Cross-Domain Human-Centered Benchmark in Monocular Depth Estimation

Model ReleasesDGX agent

arXiv:2512.08163v2 Announce Type: replace Abstract: Deep neural networks (DNNs) are increasingly used as functional models of human vision, yet standard monocular depth estimation (MDE) benchmarks lar

AdaMTP: An Adaptive Training Paradigm for Multi-Token Prediction

Model ReleasesDGX agent

arXiv:2608.00434v1 Announce Type: new Abstract: Multi-Token Prediction (MTP) has emerged as an effective paradigm that augments a shared Large Language Model backbone with auxiliary heads, training th

Adaptive Quantum Physics-Informed Neural Networks for Differential Equations with Applications to Fluid Dynamics

Model ReleasesDGX agent

arXiv:2608.00850v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have emerged as a versatile approach for solving nonlinear partial differential equations (PDEs), yet achieving

Adaptive Reconstruction of Bosonic Quantum States

Model ReleasesDGX agent

arXiv:2608.02049v1 Announce Type: cross Abstract: Bosonic quantum systems provide a hardware-efficient platform for quantum information processing but remain challenging to characterise due to their l

AdaThink-Med: Optimizing Inference-Time Compute for Medical Reasoning via Uncertainty Quantification

Model ReleasesDGX agent

arXiv:2509.24560v2 Announce Type: replace Abstract: Extended Chain-of-Thought (CoT) reasoning has significantly bolstered the capabilities of medical large language models (LLMs). However, current mod

AdvPlan-Bench: Adversarial Evaluation of Structured Plan-Generation Agents

Model ReleasesDGX agent

arXiv:2608.00832v1 Announce Type: new Abstract: Structured plan-generation agents are often evaluated as if a plan has quality in isolation, yet many realistic planning tasks require asking how a cand

AeroLLE: Constrained Pseudo-Supervision for Nighttime Aerial Image Enhancement with the AeroNight-1.5K Benchmark

Model ReleasesDGX agent

arXiv:2608.00702v1 Announce Type: new Abstract: Nighttime aerial image enhancement is challenged by spatially nonuniform exposure, mixed illumination, and weak structural evidence, while registered no

Agentic Coding in the Wild: Characterizing GitHub Copilot Traces at Production Scale

Model ReleasesDGX agent

arXiv:2608.00101v1 Announce Type: cross Abstract: AI coding agents like GitHub Copilot, Claude Code, and Codex interleave multi-step LLM inference with tool execution, creating a workload different fr

AgentMemBench: A Systematic Benchmark for Evaluating Long-Term Memory Management Strategies in Conversational AI Agents

Model ReleasesDGX agent

arXiv:2608.00009v1 Announce Type: new Abstract: Long-term memory remains a critical bottleneck for conversational AI agents, whose finite context windows cannot support coherent recall across thousand

All models are currently 20% discounted in Portal, other than GPT-5.6 Terra and Luna which 50% off and DeepSeek V4 Flash 0731 which is 90% o…

Model ReleasesDGX agent

All models in the Portal are discounted by 20 %, except GPT‑5.6 Terra and Luna (50 % off) and DeepSeek V4 Flash 0731 (90 % off). The latest Alibaba Qwen release, Qwen3.8‑Max, is now available for Herm

Amortized Inference of Multi-Modal Posteriors using Likelihood-Weighted Normalizing Flows

Model ReleasesDGX agent

arXiv:2512.04954v3 Announce Type: replace Abstract: We present a novel technique for amortized posterior estimation using Normalizing Flows trained with likelihood-weighted importance sampling. This a

An Accessible Solution for Deformable Image Registration Compared with Learning-Based Approaches

Model ReleasesDGX agent

arXiv:2608.02248v1 Announce Type: cross Abstract: Deformable image registration (DIR) is a core problem in medical image analysis; but, unlike labeling decision problems such as classification and seg

An Embedded RISC-V Evaluation of Kolmogorov--Arnold Networks in Hard-Constrained Recurrent Physics-Informed Models

Model ReleasesDGX agent

arXiv:2608.00737v1 Announce Type: new Abstract: Hard-constrained recurrent physics-informed networks (HRPINNs) embed known dynamics inside a recurrent numerical integrator and restrict a neural branch

Analyzing Speech Condition Effects in Dysarthric ASR: A Layer-wise Probing Study

Model ReleasesDGX agent

arXiv:2608.01865v1 Announce Type: new Abstract: Automatic speech recognition (ASR) performance degrades sharply on dysarthric speech, yet how disordered articulation reshapes a model's internal repres

AOSpec: Action and Observation Co-Speculation for Low-Latency Agent Serving

Model ReleasesDGX agent

arXiv:2608.00881v1 Announce Type: new Abstract: Large language model agents increasingly act through stateful tools, yet model generation and environment execution remain serialized at every step. As

Appreciate it! High performance, low cost. Try it out today. 💻

Model ReleasesDGX agent

Appreciate it! High performance, low cost. Try it out today. 💻 Qwen 3.8 Max is available on AI Gateway. It scores mid-pack on the DeepsecBench cybersecurity leaderboard at one of the lowest costs per

← Previous
1…3233343536…372
Next →