AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlog
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,597 results
Research

Rethinking Data Mixing from the Perspective of Large Language Models

DGX agent

arXiv:2604.07963v1 Announce Type: new Abstract: Data mixing strategy is essential for large language model (LLM) training. Empirical evidence shows that inappropriate strategies can significantly redu

researcharxiv-cs-cl
10 Apr 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Robustness Risk of Conversational Retrieval: Identifying and Mitigating Noise Sensitivity in Qwen3-Embedding Model

DGX agent

arXiv:2604.06176v1 Announce Type: cross Abstract: We present an empirical study of embedding-based retrieval under realistic conversational settings, where queries are short, dialogue-like, and weakly

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Stacked from One: Multi-Scale Self-Injection for Context Window Extension

DGX agent

arXiv:2603.04759v2 Announce Type: replace Abstract: The limited context window of contemporary large language models (LLMs) remains a primary bottleneck for their broader application across diverse do

model-releasesarxiv-cs-cl
10 Apr 2026
Research

Weight Group-wise Post-Training Quantization for Medical Foundation Model

DGX agent

arXiv:2604.07674v1 Announce Type: new Abstract: Foundation models have achieved remarkable results in medical image analysis. However, its large network architecture and high computational complexity

researcharxiv-cs-cv
10 Apr 2026
Model Releases

When to Call an Apple Red: Humans Follow Introspective Rules, VLMs Don't

DGX agent

arXiv:2604.06422v1 Announce Type: cross Abstract: Understanding when Vision-Language Models (VLMs) will behave unexpectedly, whether models can reliably predict their own behavior, and if models adher

model-releasesarxiv-cs-ai
10 Apr 2026
Agents

🚀Deep Agents deploy Today we’re launching Deep Agents deploy in beta. Deep Agents deploy is the fastest way to deploy a model agnostic, ope…

DGX agent

🚀Deep Agents deploy Today we’re launching Deep Agents deploy in beta. Deep Agents deploy is the fastest way to deploy a model agnostic, open source agent harness in a production ready way. Open harnes

agentsharrison-chase--x
9 Apr 2026
Model Releases

Did Mythos leak Claude Code

DGX agent

Anthropic's unreleased frontier model, Claude Mythos (internally codenamed 'Capybara'), was exposed through two separate leaks within the same week in late March/early April 2026. The Mythos model...

model-releasesemad-mostaque--x
8 Apr 2026
Industry

Sources: Stripe has finalized a deal to acquire AI model marketplace OpenRouter for more than 7B; OpenRouter was valued at 1.3B in May (Bloomberg)

DGX agent

Bloomberg: Sources: Stripe has finalized a deal to acquire AI model marketplace OpenRouter for more than 7B; OpenRouter was valued at 1.3B in May — Stripe Inc. has finalized an agreement to acquire Op

industrytechmeme
16 Aug 2026
Industry

Stripe reportedly finalizes deal to buy AI model router OpenRouter for more than $7B

DGX agent

OpenRouter Inc., which sells developers a single door into more than 400 artificial intelligence models, has agreed to terms to sell itself to Stripe Inc. for north of 7 billion, according to a report

industrysiliconangle
16 Aug 2026
Tools

Yutori's browser-use agents run in tight loops: screenshot, action, repeat, dozens of times per task. On Together AI, their Navigator model …

DGX agent

Yutori's browser-use agents run in tight loops: screenshot, action, repeat, dozens of times per task. On Together AI, their Navigator model beats frontier performance at 2x faster inference and 4-5x l

toolstogether-ai--x
16 Aug 2026
Agents

AaLLM: An End-to-End Analog Circuit Design Framework from Topology Generation to Sizing Using Large Language Models

DGX agent

arXiv:2608.13472v1 Announce Type: cross Abstract: Analog circuit design is a time-consuming, iterative process in a nonlinear and high-dimensional design space that relies heavily on expert intuition.

agentsarxiv-cs-ai
14 Aug 2026
Model Releases

Agreement Is Not Alignment: Divergent Moral Grounds in Human and LLM Ethical Judgments

DGX agent

arXiv:2608.12368v1 Announce Type: new Abstract: Agreement with human judgments is a common proxy for evaluating the alignment of large language models (LLMs). Yet agreement in final labels does not sh

model-releasesarxiv-cs-ai
14 Aug 2026
Research

CASA: Content-Acoustic Speaking Assessment with Speech Encoder and Large Language Model

DGX agent

arXiv:2608.13101v1 Announce Type: new Abstract: Research on automatic speaking assessment (ASA) has increasingly adopted multimodal speech large language models to assess learners' speaking performanc

researcharxiv-cs-cl
14 Aug 2026
Industry

CONFIRMED: @c_valenzuelab, cofounder and co-CEO of @runwayml, is speaking at Thesis, our inaugural conference on work and AI. Language model…

DGX agent

CONFIRMED: @c_valenzuelab, cofounder and co-CEO of @runwayml, is speaking at Thesis, our inaugural conference on work and AI. Language models predict the next token, aka the next unit of text. Runway’

industrycristobal-valenzuela--x
14 Aug 2026
Safety

Into the ORBIT for Time Series: Training Regimes for Foundation Models

DGX agent

arXiv:2608.13262v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have advanced primarily through architectural innovation, while training regimes for large-scale heterogeneous c

safetyarxiv-cs-ai
14 Aug 2026
Local Ai

Open models are improving at an unprecedented rate. Congrats to the whole @Zai_org team. Ollama will support GLM-5.3 upon the open release! …

DGX agent

Open models are improving at an unprecedented rate. Congrats to the whole @Zai_org team. Ollama will support GLM-5.3 upon the open release! ❤️ Get ready to code. Introducing GLM-5.3: Built to Code. Re

local-aiollama--x
14 Aug 2026
Applications

Reduced Order Modeling for Tsunami Forecasting with Bayesian Hierarchical Pooling

DGX agent

arXiv:2512.19804v2 Announce Type: replace Abstract: Reduced-order models (ROMs) can represent spatiotemporal processes in significantly fewer dimensions and can often be solved many orders of magnitud

applicationsarxiv-cs-lg
14 Aug 2026
Model Releases

Analysis of Federated Aggregation under Model Poisoning and Backdoor Attacks: A Reconstructed Cross-Dataset and Cross-Architecture Benchmark

DGX agent

arXiv:2608.11423v1 Announce Type: cross Abstract: Robust comparisons of federated aggregation methods require joint consideration of predictive performance, threat definitions, metric semantics, and e

model-releasesarxiv-cs-cv
13 Aug 2026
Research

Boundary-Continuous Cross-Camera RGB Mapping via Hue-Split Model Trees

DGX agent

arXiv:2608.11548v1 Announce Type: cross Abstract: We propose a hue-split model-tree method for boundary-continuous cross-camera RGB mapping. Cross-camera RGB mapping aims to produce consistent color r

researcharxiv-cs-cv
13 Aug 2026
Safety

CLAIM: Leading Open-domain Active Clarification of Large Language Models with Uncertainty Measurement

DGX agent

arXiv:2608.11631v1 Announce Type: new Abstract: In open-domain human-computer interaction scenarios, large language models (LLMs) frequently encounter user queries that are ambiguous or incomplete. In

safetyarxiv-cs-ai
13 Aug 2026
Research

CORA-Diff: Confidence-Oriented Residual Acceptance for Efficient Diffusion Language Model Inference

DGX agent

arXiv:2608.11235v1 Announce Type: new Abstract: Diffusion language models (DLMs) update many tokens in parallel, yet practical decoders often use a fixed denoising horizon. Many predictions stabilize

researcharxiv-cs-ai
13 Aug 2026
Local Ai

Distillation of Foundation Models for Time-dependent PDEs

DGX agent

arXiv:2608.11937v1 Announce Type: new Abstract: Foundation models for time-dependent partial differential equations (PDEs) are trained on large and diverse collections of physical systems and can gene

local-aiarxiv-cs-lg
13 Aug 2026
Research

Hamilton-Zero: A Neural Tensor-Network Foundation Model for Ground States of Arbitrary Quadratic Qubit Hamiltonians

DGX agent

arXiv:2608.11911v1 Announce Type: cross Abstract: A central promise of useful quantum advantage is the ability to compute ground states of Hamiltonian systems beyond the reach of classical simulation

researcharxiv-cs-ai
13 Aug 2026
Agents

I think this will turn out to be wrong, and not just because I suspect there are greater returns to more intelligent models than people expe…

DGX agent

I think this will turn out to be wrong, and not just because I suspect there are greater returns to more intelligent models than people expect Economic value comes from agents, not chatbots. And accur

agentsethan-mollick--x
13 Aug 2026
Research

Large Language Model-Driven Small-Capitalization Trading: Integrating Financial News Sentiment, Macroeconomic Indicators, and Technical Signals

DGX agent

arXiv:2608.12283v1 Announce Type: cross Abstract: Large language models can extract richer signals from financial news than fixed sentiment lexicons, and recent work has explored feeding such signals

researcharxiv-cs-cl
13 Aug 2026
Safety

LazyTrain: Limited-resource Allocation toward Zero-waste Yield Optimization in Large Language Model Training

DGX agent

arXiv:2608.11919v1 Announce Type: new Abstract: Training large language models on limited hardware is increasingly a scheduling problem across GPU compute, host memory, PCIe transfer, and storage band

safetyarxiv-cs-cl
13 Aug 2026
Applications

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation

DGX agent

arXiv:2606.17598v2 Announce Type: replace-cross Abstract: Humans naturally leverage diverse sensing modalities to interact with the physical world, while most Vision-Language-Action (VLA) models for r

applicationsarxiv-cs-cv
13 Aug 2026
Model Releases

OEIS Open: How many conjectures can language models turn into theorems?

DGX agent

arXiv:2608.11941v1 Announce Type: new Abstract: We construct OEIS Open, a benchmark based on 492 open mathematical conjectures from the OEIS, formalized in Lean by Tsoukalas et al. Whereas these conje

model-releasesarxiv-cs-ai
13 Aug 2026
Safety

RA-ClipScore: Making Generative Model Evaluation More Interpretable

DGX agent

arXiv:2608.12088v1 Announce Type: new Abstract: Generative models can produce images nearly indistinguishable from real data, yet rigorous and interpretable evaluation remains challenging. Conventiona

safetyarxiv-cs-cv
13 Aug 2026
Research

RLinf-VLA: A Unified and Efficient Framework for Reinforcement Learning of Vision-Language-Action Models

DGX agent

arXiv:2510.06710v3 Announce Type: replace Abstract: Recent studies have demonstrated the potential of reinforcement learning (RL) to improve the task performance of vision-language-action (VLA) models

researcharxiv-cs-ro
13 Aug 2026
Safety

Sources: Demis Hassabis pitched a new independent industry AI safety entity, modeled on the IAEA, to top Trump officials before stepping down as DeepMind's CEO (Wall Street Journal)

DGX agent

Wall Street Journal: Sources: Demis Hassabis pitched a new independent industry AI safety entity, modeled on the IAEA, to top Trump officials before stepping down as DeepMind's CEO — Demis Hassabis di

safetytechmeme
13 Aug 2026
Research

TELLME: Test-Enhanced Learning for Language Model Enrichment

DGX agent

arXiv:2608.11788v1 Announce Type: cross Abstract: Continual pre-training (CPT) has been widely adopted as a method for domain adaptation in large language models. However, CPT has consistently been ac

researcharxiv-cs-ai
13 Aug 2026
Safety

Towards Human Motion World Models via Executable Behaviour Representations

DGX agent

arXiv:2604.18064v2 Announce Type: replace Abstract: Human motion world models should capture motion's intentionality by being executable: adaptable to different actions and capable of assessing motion

safetyarxiv-cs-ai
13 Aug 2026
Model Releases

TRACES: A Benchmark for Epistemic Reliability in Scientific Reasoning by LLMs

DGX agent

arXiv:2608.11415v1 Announce Type: cross Abstract: Large language models are being proposed as agents in scientific workflows, in domains where no downstream verifier exists. Such deployment assumes th

model-releasesarxiv-cs-ai
13 Aug 2026
Research

Delays in Spiking Neural Networks: A State Space Model Approach

DGX agent

arXiv:2512.01906v3 Announce Type: replace Abstract: Spiking neural networks (SNNs) are biologically inspired, event-driven models suited for temporal data processing and energy-efficient neuromorphic

researcharxiv-cs-lg
12 Aug 2026
Model Releases

DreamOmni3: Scribble-based Editing and Generation

DGX agent

arXiv:2512.22525v2 Announce Type: replace Abstract: Recently unified generation and editing models have achieved remarkable success with their impressive performance. These models rely mainly on text

model-releasesarxiv-cs-cv
12 Aug 2026
Applications

Graphical Models of False Information and Fact Checking Ecosystems

DGX agent

arXiv:2208.11582v2 Announce Type: replace-cross Abstract: The wide spread of false information online, including misinformation and disinformation, has become a major problem for our highly digitised

applicationsarxiv-cs-ai
12 Aug 2026
Research

Hybrid Token Compression for Vision-Language Models

DGX agent

arXiv:2512.08240v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) rely on hundreds of visual tokens, leading to high computational and memory costs. Existing compression methods

researcharxiv-cs-ai
12 Aug 2026
Model Releases

I ran DeepSeek V4 Flash 284B + DSpark on one RTX PRO 6000. The drafter was faster in RAM than VRAM.

DGX agent

Hey guys, Just finished benchmarking DeepSeek V4 Flash 284B + DSpark on a single RTX PRO 6000 96GB. Short version: DSpark: ~15–17% faster generation on my coding workload On this setup, the DSpark dra

model-releasesr-localllama
12 Aug 2026
Research

Logit Lens Supervision for Patch-Level Explanations in Vision-Language Models

DGX agent

arXiv:2602.01530v2 Announce Type: replace Abstract: Modern autoregressive Vision-Language Models (VLMs) can generate fluent answers while their visual-token representations become weakly tied to the i

researcharxiv-cs-cv
12 Aug 2026
Safety

MarkNull: Model-Agnostic Watermark Removal in AI-Generated Images via On-Manifold Latent Manipulation

DGX agent

arXiv:2608.10166v1 Announce Type: cross Abstract: Digital watermarking has emerged as a critical technique for provenance and copyright attribution in AI-generated imagery, yet its robustness against

safetyarxiv-cs-ai
12 Aug 2026
Model Releases

Multimodal Item Parameter Estimation using Simulated Response Probabilitie

DGX agent

arXiv:2608.10154v1 Announce Type: cross Abstract: We present results from reconstructing multiple-choice model (MCM) and three-parameter logistic (3PL) model curves using a fine-tuned multimodal large

model-releasesarxiv-cs-ai
12 Aug 2026
Safety

OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation

DGX agent

arXiv:2603.19201v3 Announce Type: replace Abstract: Contact-rich manipulation tasks, such as wiping and assembly, require accurate perception of contact forces, friction changes, and state transitions

safetyarxiv-cs-ro
12 Aug 2026
Agents

On Understanding, Identifying, and Mitigating Vulnerabilities in Agentic Large Language Models

DGX agent

arXiv:2608.10530v1 Announce Type: cross Abstract: Large Language Models (LLMs) have undergone a shift from stateless conversational interfaces to autonomous agents capable of multi-step planning, tool

agentsarxiv-cs-ai
12 Aug 2026
Research

VIScore: Diagnosing Planning-Relevant Quality in Latent World Models

DGX agent

arXiv:2608.11174v1 Announce Type: new Abstract: Regulating the latent space to an isotropic Gaussian distribution provides a stable and information-maximized landscape for world model planning. Howeve

researcharxiv-cs-ro
12 Aug 2026
Model Releases

AgriField-40K: Adapting Vision Models to Agriculture With Efficient Continual Pretraining

DGX agent

arXiv:2608.07984v1 Announce Type: new Abstract: Field-based agricultural computer vision is important for precision agriculture, yet it largely depends on expensive annotations and costly adaptation o

model-releasesarxiv-cs-cv
11 Aug 2026
Local Ai

Auditing Medical Vision-Language Models on Chest Radiographs: Estimating Reference Agreement Across Institutions

DGX agent

arXiv:2608.07550v1 Announce Type: new Abstract: Vision-language models return structured chest-radiograph findings through interfaces exposing no confidence score, so a receiving institution cannot re

local-aiarxiv-cs-cv
11 Aug 2026
Research

Commitment Before Realization: When Classifier-Free Guidance Becomes Unnecessary in Masked Diffusion Language Models

DGX agent

arXiv:2608.08082v1 Announce Type: new Abstract: Classifier-free guidance (CFG) is usually kept on throughout masked diffusion language model decoding, although its benefit varies across prompts and ov

researcharxiv-cs-cl
11 Aug 2026
← Previous
1…182183184185186…1263
Next →