AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,987 results
18 Aug 2026

🚀Today we launched LangSmith Tuned Evaluators, starting with Perceived Error. Tuned Evaluators run on production traces to catch undesirabl…

Model ReleasesDGX agent

🚀Today we launched LangSmith Tuned Evaluators, starting with Perceived Error. Tuned Evaluators run on production traces to catch undesirable agent behavior and attach feedback that you can use in your

TransAnyText: Translating Arbitrary Text in E-commerce Images via Structured Visual Generation

Model ReleasesDGX agent

arXiv:2608.16284v1 Announce Type: new Abstract: Cross-border e-commerce image translation is essential for global retail, where product images, banners, and detail pages need to be produced in differe

We have paused some frontier RL training to ensure that we can meet the appropriate alignment, security and monitoring standards for the new…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety
DGX agent

We have paused some frontier RL training to ensure that we can meet the appropriate alignment, security and monitoring standards for the new level of capabilities in front of us. Model progress is now

When Do LLMs Apply the Wrong Law? Diagnosing LLM Failures in Temporal Legal Reasoning

Model ReleasesDGX agent

arXiv:2608.14610v1 Announce Type: new Abstract: Legal reasoning tasks such as legal judgment prediction (LJP) require identifying the temporally correct version of the law governing a case -- a capabi

17 Aug 2026

Act2Intention: A Benchmark For Developing Active Mobile Agents Through Inferring User Intention from GUI Actions

Model ReleasesDGX agent

arXiv:2608.14132v1 Announce Type: cross Abstract: Mobile GUI Agents powered by multimodal large language models (MLLMs) show promise in human-computer intelligence. However, current research primarily

Adversarial Learning of Classifier-Free Guidance Schedules

SafetyDGX agent

arXiv:2608.14038v1 Announce Type: new Abstract: Modern text-to-image diffusion models rely on classifier-free guidance (CFG) to achieve high image fidelity and text alignment. However, CFG typically a

Anthropic explains how Claude’s invisible text watermarks will work

Model ReleasesDGX agent

Anthropic has clarified how it's planning to apply invisible watermarks to Claude-generated text in order to comply with Europe's AI transparency rules. On Friday, Anthropic announced that Claude's te

ARC: Fair Relative Advantage Comparison in Open-Ended Real-World Interaction

Model ReleasesDGX agent

arXiv:2608.13622v1 Announce Type: new Abstract: Open-ended real-world interaction admits multiple valid behaviors: an agent may answer directly, ask for clarification, provide progress updates, or con

BCMT: Blockwise Causal Memory Transformer

ResearchDGX agent

arXiv:2608.13578v1 Announce Type: cross Abstract: Transformer architectures rely on dense self-attention to model long-range dependencies, but this mechanism exhibits quadratic complexity with respect

BM25-Augmented Many-Shot Translation for Low-Resource North-Eastern Indian Languages

Model ReleasesDGX agent

arXiv:2608.13722v1 Announce Type: new Abstract: This paper describes the University of Florida Gators submission to the WMT26 Low-Resource Indic Language Translation shared task. We adapt the retrieva

Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination

Model ReleasesDGX agent

arXiv:2608.14391v1 Announce Type: cross Abstract: Recent video generators can fabricate realistic depictions of wars, disasters, public emergencies, and other real-world crises, creating substantial r

Capacity-Dependent Effects of Data Selection for Reasoning

ResearchDGX agent

arXiv:2608.13721v1 Announce Type: cross Abstract: In reasoning supervised fine-tuning, candidate responses for the same instruction can differ substantially in how well they match the student's curren

🆕 Context Engineering in 2026: Compaction, Memory & Cost https://www.youtube.com/watch?v=WP3hjUXd918 @Whats_AI, @samridhivaid and @omar_sol…

Model ReleasesDGX agent

🆕 Context Engineering in 2026: Compaction, Memory & Cost https://www.youtube.com/watch?v=WP3hjUXd918 @Whats_AI, @samridhivaid and @omar_solano1 return! This workshop is about engineering the context w

Does a Language Server Save Tokens for Coding Agents? A Measurement Methodology and Preliminary Study

Model ReleasesDGX agent

arXiv:2608.13568v1 Announce Type: cross Abstract: Coding agents spend most of their context budget on retrieval. Lexical retrieval (grep) is universal, instant, and zero-setup, but noisy: it cannot te

Federated Prompt Learning: A Unified Framework, Empirical Analysis, and Future Directions

ResearchDGX agent

arXiv:2608.13844v1 Announce Type: cross Abstract: Large language models (LLMs) have become core components of cloud-based intelligent services in academia and industry, yet their training and deployme

How accurate do you think this is? Qwen3.5 9B vs GPT-4o

Model ReleasesDGX agent

Do you think today’s GPU-poor systems running Qwen 3.5 9B can outperform the ones we had with ChatGPT-4o? I mean, when 4o disappeared, people felt like they’d lost a great model, and now Qwen 3.5 9B Q

Implementing Computational Law in Wolfram Language for the Governance of Artificial Intelligence

Model ReleasesDGX agent

arXiv:2608.13958v1 Announce Type: new Abstract: How do we govern AI systems whose reasoning we cannot fully inspect? Governance does not require understanding a system's reasoning. It requires stating

Language-Specific Gaps in AI Safety Training Datasets

SafetyDGX agent

arXiv:2608.13695v1 Announce Type: cross Abstract: Large language model providers routinely cite multilingual safety benchmarks spanning a dozen or more languages as evidence that their models are safe

Learning-to-Transition for Large-scale and High-Order MIMO Detection

Model ReleasesDGX agent

arXiv:2608.14511v1 Announce Type: cross Abstract: High-order multiple-input multiple-output (MIMO) detection requires efficient search over a large discrete symbol space while producing reliable soft

MAGneT-3D: Monocular and Domain-Generalizable Temporal 3D Detection

Model ReleasesDGX agent

arXiv:2608.14282v1 Announce Type: new Abstract: Monocular temporal 3D detection aims to detect objects in 3D, given a monocular video. Query-based 3D detectors unify detection and cross-view associati

Omni-LiveAvatar: Minute-Level Real-Time Streaming Joint Audio-Visual Avatar Generation

HardwareDGX agent

arXiv:2608.13602v1 Announce Type: cross Abstract: Joint audio-video generative models serve as foundation for immersive and interactive digital-human generation. Nevertheless, most existing models rel

Owner3D: Ownership-Guided Style Writing for Training-Free Localized 3D Stylization

Model ReleasesDGX agent

arXiv:2608.14078v1 Announce Type: new Abstract: Localized 3D stylization aims to modify the appearance of a specified object part while preserving the remaining surfaces. In large reconstruction model

READ: A Retrieval-Alignment Diffusion Framework for Structure-based Drug Design

SafetyDGX agent

arXiv:2506.14488v2 Announce Type: replace-cross Abstract: Structure-based drug design (SBDD) models are central to modern pharmaceutical research, enabling the rational exploration of protein-ligand i

Regime-Conditional Verification: Correctness Estimation for Adapting and Monitoring Safety Classifiers

Model ReleasesDGX agent

arXiv:2608.14089v1 Announce Type: new Abstract: Safety classifiers deployed with large language models often fail for two reasons: their decisions reflect the policy learned during training rather tha

Retrieve-then-Adapt: Retrieval-Augmented Test-Time Adaptation for Sequential Recommendation

Model ReleasesDGX agent

arXiv:2604.05379v2 Announce Type: replace-cross Abstract: The sequential recommendation (SR) task aims to predict the next item based on users' historical interaction sequences. Typically trained on h

RGBX-Next: Towards Realistic Generative Rendering from G-Buffers

ResearchDGX agent

arXiv:2608.13929v1 Announce Type: new Abstract: Diffusion models have achieved impressive results in image, video, and streaming generation. However, compared to traditional 3D rendering, they still l

SAGE: Surrogate-gradient Adaptation via Attention-Guided Entropy for Spiking Transformers

Model ReleasesDGX agent

arXiv:2608.13702v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) offer an energy-efficient alternative to conventional deep neural networks by exploiting sparse event-driven computatio

Simulation-Aware In-Context Policy Improvement for LLM-Aided Analog Layout Refinement

Model ReleasesDGX agent

arXiv:2608.13767v1 Announce Type: new Abstract: Analog IC layout design remains a labor-intensive iterative process dominated by simulation-driven refinement. Although end-to-end layout generators acc

Simulation-Driven Vehicular Traffic Data Augmentation: Extending Sensor Coverage Through Virtual Sensing

ResearchDGX agent

arXiv:2608.13993v1 Announce Type: new Abstract: Urban traffic management relies on sensor networks whose spatial coverage is limited by deployment costs and privacy regulations. Machine learning model

The Dynamics of Intelligence Explosions

Model ReleasesDGX agent

arXiv:2608.14426v1 Announce Type: new Abstract: AI is increasingly being used to help with AI R&D. Under certain conditions this feedback loop might be able to produce an intelligence explosion, with

XtraLight-MedMamba for Classification of Neoplastic Tubular Adenomas

Model ReleasesDGX agent

arXiv:2602.04819v5 Announce Type: replace Abstract: Accurate risk stratification of precancerous polyps during routine colonoscopy screening is a key strategy to reduce the incidence of colorectal can

Zero-Shot Skeleton-Based Action Anticipation

Model ReleasesDGX agent

arXiv:2608.14243v1 Announce Type: new Abstract: Action anticipation (AA) aims to recognize ongoing human or humanoids actions from partial observations, enabling robots to predict intentions before th

16 Aug 2026

Anyone managed to get Qwen 3.8 27B running smoothly on vLLM? Can't get rid of endless thinking

Model ReleasesDGX agent

Title pretty much says it all. I’ve deployed Qwen 3.8 27B using vLLM on an RTX 6000 Pro (tried multiple vLLM releases and launch recipes), but I can't get it into a usable state because of crazy long

b10453

Model ReleasesDGX agent

model : remove some ggml_concat (#27176) Co-authored-by: Xuan Son Nguyen son@huggingface.co Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabl

Best Setup for a 16 GB VRAM + 128 GB RAM System?

Model ReleasesDGX agent

Running a 12700k + 5060 Ti 16 gb with 128 gb DDR4 ram and I'm wondering what's the ideal setup to maximize performance. I did some preliminary stuff but I have to admit, I'm still learning and kinda j

Huihui-ai Qwen 3.8 Ablit Available

Model ReleasesDGX agent

At hugging face, this is the model I use for most of my analysis that most of my services have to not be refused. previous versions are quite good. It seems this might have dropped today and pulling r

Mindblow with Qwen 3.8 (Cline + VSC on a 5090 mobile)

Model ReleasesDGX agent

Honestly I am surprised at the speed and performance I am getting from the model, it does not allow for complete hands off like Claude 5.0 but that is not what I want, i want to be able to iterate and

We tuned an AI agent that can do large-scale document extraction from long docs (50+ pages, some with 10k-100k fields) with 94%+ accuracy 📈…

Model ReleasesDGX agent

We tuned an AI agent that can do large-scale document extraction from long docs (50+ pages, some with 10k-100k fields) with 94%+ accuracy 📈 It uses a harness + model set that is tuned specifically for

15 Aug 2026

1/2 Thanks Gavin for an especially thoughtful exchange. I don't usually spend much time on social media but I wanted to engage here because …

SafetyDGX agent

1/2 Thanks Gavin for an especially thoughtful exchange. I don't usually spend much time on social media but I wanted to engage here because it really brings out the heart of an important conversation.

b10447

Model ReleasesDGX agent

server: re-design yield_to_queue thread model (#27133) run common_speculative_process in worker swap worker <--> main thread design Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) ma

14 Aug 2026

1BIT Qwen 3.8 2.4T a95b (unsloth iQ1_S) (MEDIUM Reasoning)

Model ReleasesDGX agent

Processing img az99qopcg8jh1... So same as my prior post 1bit test... although this 1bit is a bit interesting you can read on unlsoth blog https://unsloth.ai/docs/models/qwen3.8 508 gigs being used I

A Contract-Grade Verifier for LLM-Generated GPU Kernels, and a Native Blackwell Backward for the Gated-Linear-Recurrence Family

Model ReleasesDGX agent

arXiv:2608.12700v1 Announce Type: new Abstract: Systems that generate GPU kernels with language models report high correctness rates. Those rates come from a single loose test: run the kernel on a few

A Unified Framework for Joint Detection of Lacunes and Enlarged Perivascular Spaces

Model ReleasesDGX agent

arXiv:2603.04243v3 Announce Type: replace Abstract: Cerebral small vessel disease (CSVD) markers, specifically enlarged perivascular spaces (EPVS) and lacunae, present a unique challenge in medical im

b10434

Model ReleasesDGX agent

chat : pass reasoning_effort to template chat: add reasoning_effort to common_chat_templates_inputs Store OpenAI Chat Completions reasoning_effort and make it available to jinja templates (with model

Balanced Adaptive Prototype Selection for Scalable TabPFN Inference on Large-Scale Tabular Data

HardwareDGX agent

arXiv:2608.12989v1 Announce Type: new Abstract: Pretrained tabular foundation models have demonstrated strong predictive capability; however, their application to large-scale datasets remains constrai

CROP: Task Relevance via Counterfactuals for Selective On-Policy Distillation

SafetyDGX agent

arXiv:2608.13387v1 Announce Type: new Abstract: On-policy distillation (OPD) supervises a student language model on trajectories sampled from its current policy, but assigns equal credit to response t

damn deepseek moves fast on roon tweets

Model ReleasesDGX agent

damn deepseek moves fast on roon tweets the major ai companies should commit to a real time priced API product. ai demand varies wildly over a day/night curve and the industry is broadly extremely cap

Dual-Flow Transformers: Decoupling the Primary Prefill Path from Additional Decode Computation

ResearchDGX agent

arXiv:2608.12385v1 Announce Type: new Abstract: As large language models serve more requests, cumulative inference cost is becoming increasingly important relative to one-time training cost. The two i

Enhancing In-Hospital Mortality Prediction Using Multi-Representational Learning with LLM-Generated Expert Summaries

ResearchDGX agent

arXiv:2411.16818v2 Announce Type: replace-cross Abstract: To evaluate a multi-representational framework in which large language model (LLM)-generated expert summaries of intensive care unit (ICU) not

ERSkill: Evolving for Skill-Guided Adaptive Memory Retrieval

Model ReleasesDGX agent

arXiv:2608.12720v1 Announce Type: cross Abstract: While Large Language Model (LLM) agents increasingly rely on long-term memory for persistent interactions, the retrieval mechanisms governing this mem

Excited to see Qwen3.8 running at scale with TokenSpeed! 🚀 Light on latency, big on speed. Kudos to LightSeek for the fantastic Day-0 suppo…

Model ReleasesDGX agent

Excited to see Qwen3.8 running at scale with TokenSpeed! 🚀 Light on latency, big on speed. Kudos to LightSeek for the fantastic Day-0 support! @lightseekorg We’re proud to be the Day 0 open-source inf

Fixed/improved Jinja chat template for Qwen 3.8

Model ReleasesDGX agent

'Again?' you might ask. I(*) took the original Qwen 27B 3.8 chat template and compared it to the improved template that was posted a day before. The original has issues. The improved one fixed some wh

FUSE: Active Functional Affordance Grounding through Adaptive Semantic-Geometric Evidence Acquisition

Model ReleasesDGX agent

arXiv:2608.12683v1 Announce Type: cross Abstract: Embodied agents must often identify and interact with objects based on their function rather than their identity, requiring them to actively acquire o

GENADA: efficient generative time series adversarial attack framework

ApplicationsDGX agent

arXiv:2608.12535v1 Announce Type: new Abstract: Deep learning models are widely used for time series analysis in domains such as healthcare, finance, energy systems, and environmental monitoring. Howe

HumanoidVLN: A Physics-Grounded Simulator and Benchmark for Vision-Language Navigation Across Diverse Humanoid Embodiments

Model ReleasesDGX agent

arXiv:2608.12860v1 Announce Type: new Abstract: Vision-Language Navigation (VLN) for humanoid robots poses challenges existing benchmarks fail to address: bipedal locomotion imposes physical constrain

HybridSB-MoE: Dual-Domain Schrodinger Bridges with Scene-Adaptive Expert Routing for Speech Enhancement

SafetyDGX agent

arXiv:2608.12715v1 Announce Type: cross Abstract: Generative speech enhancement faces three gaps: spectral models capture harmonic structure but often disrupt phase, waveform models preserve phase but

LOB-ID: Evaluating Synthetic Market Data by Inception Distances

ResearchDGX agent

arXiv:2608.13082v1 Announce Type: cross Abstract: Generative models of limit orderbook (LOB) data have advanced rapidly, but their evaluation often focuses on stylised facts and selected market statis

MAG: MAnifold Guided Semi-Supervised Multi-modal In-Context Learning

Model ReleasesDGX agent

arXiv:2608.12724v1 Announce Type: new Abstract: Few-shot in-context learning (ICL) with multi-modal large language models (MLLMs) enables task adaptation without parameter updates, but its performance

Mistral is now hosting GLM-5.2

Model ReleasesDGX agent

Not directly LOCALLlama related but I thought it was interesting since Mistral and Z.ai are competitors, and more surprisingly they are pricing it (GLM-5.2) even cheaper than their current flagship mo

MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classification

AgentsDGX agent

arXiv:2608.13463v1 Announce Type: cross Abstract: Modern image classification models excel when trained on single task-specific datasets but often struggle to generalize across domains and difficulty

← Previous
1…374375376377378…1050
Next →