AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Safety

SEMA: Simple yet Effective Learning for Multi-Turn Jailbreak Attacks

DGX agent

arXiv:2602.06854v2 Announce Type: replace Abstract: Multi-turn jailbreaks capture the real threat model for safety-aligned chatbots, where single-turn attacks are merely a special case. Yet existing a

safetyarxiv-cs-cl
14 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Semantic Steering for Controllable Generation: Tuning-Free Concept Erasure in Multimodal Diffusion Transformers

DGX agent

arXiv:2608.12829v1 Announce Type: new Abstract: Multimodal Diffusion Transformers (MM-DiTs) have demonstrated remarkable text-to-image generation performance, surpassing traditional U-Net-based diffus

safetyarxiv-cs-cv
14 Aug 2026
Safety

SpatialVAM:Spatial-Aware Multi-View Video Diffusion as a Data-Efficient Robot Policy

DGX agent

arXiv:2604.03181v2 Announce Type: replace-cross Abstract: Robotic manipulation requires understanding both the 3D spatial structure of the environment and its temporal evolution, yet most existing pol

safetyarxiv-cs-cv
14 Aug 2026
Safety

StateBridge: Training-free Hidden-state Alignment for Latent Communication in LLM Multi-Agent Systems

DGX agent

arXiv:2608.13317v1 Announce Type: new Abstract: Large language model based multi-agent systems usually communicate in text, i.e., using discrete tokens. However, text introduces a discrete bottleneck.

safetyarxiv-cs-ai
14 Aug 2026
Model Releases

StreamTTT: Reconciling Real-Time Perception and Long-Term Memory in Streaming VLMs

DGX agent

arXiv:2608.13416v1 Announce Type: new Abstract: Humans effortlessly perceive the present while remembering the past, yet streaming VLMs often trade off real-time perception against long-term memory. P

model-releasesarxiv-cs-cv
14 Aug 2026
Safety

Synthetic Persona Pretraining: Alignment from Token Zero

DGX agent

arXiv:2608.13482v1 Announce Type: cross Abstract: As language-model-based AI is increasingly deployed in autonomous settings, aligning its goals and values with those of humans becomes critical. Today

safetyarxiv-cs-ai
14 Aug 2026
Safety

Unmasking Conversational Bias in AI Multiagent Systems

DGX agent

arXiv:2501.14844v3 Announce Type: replace-cross Abstract: Detecting biases in the outputs produced by generative models is essential to reduce the potential risks associated with their application in

safetyarxiv-cs-ai
14 Aug 2026
Research

V-RAE: Rethinking Video Latent Spaces for Generation

DGX agent

arXiv:2608.13556v1 Announce Type: new Abstract: Latent video generation relies on autoencoders to define a compact space in which generative models operate. Although video autoencoder architectures ha

researcharxiv-cs-cv
14 Aug 2026
Research

Adaptive Online Learning with LSTM Networks for Energy Price Prediction

DGX agent

arXiv:2510.16898v2 Announce Type: replace-cross Abstract: Accurate prediction of electricity prices is crucial for stakeholders in the energy market, particularly for grid operators, energy producers,

researcharxiv-cs-ai
13 Aug 2026
Model Releases

b10415

DGX agent

spec : auto-detect mtp draft model type (#27005) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramew

model-releasesllama-cpp-releases
13 Aug 2026
Model Releases

BrowseSafe: Understanding and Preventing Prompt Injection Within AI Browser Agents

DGX agent

arXiv:2511.20597v2 Announce Type: replace-cross Abstract: The integration of artificial intelligence (AI) agents into web browsers introduces security challenges that go beyond traditional web applica

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Can Frontier LLMs Match Natively Multimodal Embeddings? A Comparison on Hard-Negative Text-to-Image Retrieval

DGX agent

arXiv:2608.11343v1 Announce Type: new Abstract: Multimodal retrieval and classification across different types of media, spanning text, images,video and audio, has traditionally relied on dual-encoder

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Causal Structure is Inducible but Functionally Decoupled: The Routing/Readout Boundary of a Typed Mechanism Library

DGX agent

arXiv:2608.11767v1 Announce Type: new Abstract: When a language model answers an interventional question, the computation it must perform depends on the type of evidence the query requires. We report

model-releasesarxiv-cs-cl
13 Aug 2026
Research

Certifying What Helps Customer-Return Timing: A Screen-and-Confirm Test for Conditioning Signals, and Why Decay Is Nearly Enough

DGX agent

arXiv:2608.11555v1 Announce Type: new Abstract: Practitioners enrich customer-return models with ever more signals (lifetime value, category, recency/frequency, calendar, geography), and the temporal-

researcharxiv-cs-lg
13 Aug 2026
Model Releases

Commonsense on Demand: Generating and Selectively Integrating Commonsense Knowledge for Natural Language Inference

DGX agent

arXiv:2507.15100v3 Announce Type: replace-cross Abstract: Natural Language Inference (NLI) determines whether a premise entails, contradicts, or is neutral with respect to a hypothesis. The task is of

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Diffuse to Compress: Leveraging Diffusion LMs for Lossless Compression

DGX agent

arXiv:2608.11249v1 Announce Type: cross Abstract: We study the problem of lossless text compression, motivated by the rapid growth in the collection and storage of digital textual data - including pla

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Explainability in Practice: A Survey of Explainable NLP Across Various Domains

DGX agent

arXiv:2502.00837v3 Announce Type: replace-cross Abstract: Natural Language Processing (NLP) is now embedded in critical sectors including healthcare, finance, and customer relationship management, whe

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

From Synthesis to Removal: Physics-Grounded Reflection Simulation and Diffusion-Based Video Dereflection

DGX agent

arXiv:2608.11562v1 Announce Type: cross Abstract: Videos captured through glass often contain reflections that degrade visual quality and interfere with downstream vision tasks. Although single-image

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

ggml-cpu/ops: vectorize flash-attention V-cache F16 to F32 conversion by jinzihao · Pull Request #26947 · ggml-org/llama.cpp

DGX agent

Overview ggml_cpu_fp16_to_fp32 leverages hardware F16C intrinsics (AVX-512, AVX2, etc.), faster than the software-only ggml_fp16_to_fp32_row, bringing 17-31% gain in prompt processing rate for a small

model-releasesr-localllama
13 Aug 2026
Research

HAMP-LIC: Hessian-Aware Mixed-Precision Post-Training Quantization for Learned Image Compression

DGX agent

arXiv:2608.12239v1 Announce Type: cross Abstract: Use this plain-text version for the arXiv abstract field: Learned image compression (LIC) models achieve strong rate-distortion performance but are hi

researcharxiv-cs-ai
13 Aug 2026
Model Releases

Harnessing agent memory to build lifelong AI partners for materials scientists

DGX agent

arXiv:2608.11224v1 Announce Type: new Abstract: Materials research advances through accumulated experience - scripts that work, protocols that are trusted, warnings attached to failed calculations or

model-releasesarxiv-cs-ai
13 Aug 2026
Research

How Far from Clinical Deployment? Evaluating the Complete Unsupervised Domain Adaptation Pipeline in Medical Imaging

DGX agent

arXiv:2608.12035v1 Announce Type: cross Abstract: Deploying unsupervised domain adaptation (UDA) in clinical practice requires choosing which algorithm to use and which of its trained models to ship.

researcharxiv-cs-ai
13 Aug 2026
Model Releases

Interesting uses for Muse Glimmer 30B?

DGX agent

Hey everyone, Non-native speaker, writing my post by hand, let me know if I make mistakes (can only learn from it!) Muse Glimmer 30B is so far quite nice, but I haven't found a clear-cut case yet what

model-releasesr-localllama
13 Aug 2026
Model Releases

JieZi: A Large-Scale Expert-Audited Dataset and Benchmark for Ancient Chinese Character Exegesis

DGX agent

arXiv:2608.11741v1 Announce Type: cross Abstract: The scholarly exegesis of ancient Chinese characters demands integrating visual observation, linguistic analysis, and historical context. However, exi

model-releasesarxiv-cs-ai
13 Aug 2026
Research

Learning with Bilevel-Minimax Optimization for Efficient and Reliable Transfer Attacks

DGX agent

arXiv:2608.11815v1 Announce Type: cross Abstract: Transfer-based adversarial attacks craft adversarial examples using surrogate models to mislead black-box victim models. Beyond perturbation generatio

researcharxiv-cs-cv
13 Aug 2026
Local Ai

LinearKV: One Cached State Suffices for Position-Independent Caching in Hybrid LLMs

DGX agent

arXiv:2608.11231v1 Announce Type: new Abstract: LLM serving is increasingly accelerated by position-independent caching (PIC). Existing PIC methods, however, are built for full-attention models, where

local-aiarxiv-cs-ai
13 Aug 2026
Model Releases

MBA: Multimodal Benchmark and Agents for Real-World Business Ideation

DGX agent

arXiv:2608.11616v1 Announce Type: new Abstract: Agentic systems powered by large language models (LLMs) have opened new opportunities for business ideation. Yet existing approaches remain confined to

model-releasesarxiv-cs-ai
13 Aug 2026
Local Ai

OrderMoE: An expert similarity driven distributed edge MoE inference

DGX agent

arXiv:2607.17154v2 Announce Type: replace-cross Abstract: Although mixture-of-experts, MoE, models have been increasingly adopted to scale large language models with moderate computation cost, it rema

local-aiarxiv-cs-lg
13 Aug 2026
Research

PolarSym: Polar Geometry-aware Attention for CAD Floorplan Parsing

DGX agent

arXiv:2608.11793v1 Announce Type: new Abstract: CAD plan parsing is a fundamental task in Building Information Modeling (BIM), aiming to automatically extract architectural elements including walls, d

researcharxiv-cs-cv
13 Aug 2026
Model Releases

Previewing Ultrafast mode: GPT-5.6 Sol at up to 14x the speed. Launching first in the OpenAI API to a select group of customers with expande…

DGX agent

On Aug 13 2026, OpenAI announced a preview of its “Ultrafast” mode for the GPT‑5.6 model, called Sol, which claims inference speeds up to 14× faster than current releases. The feature is initially rol

model-releasesopenai--x
13 Aug 2026
Research

Reinforcing Step-level Reasoning for Effective Self-Correction in LLMs

DGX agent

arXiv:2608.11573v1 Announce Type: cross Abstract: Achieving effective self-correction, where models verify and correct their own mistakes, remains a fundamental challenge for large language models (LL

researcharxiv-cs-ai
13 Aug 2026
Model Releases

Rubric Dropout: A Simple Way to Mitigate Reward Hacking in Rubric-as-Reward RL

DGX agent

arXiv:2608.11669v1 Announce Type: cross Abstract: Reinforcement learning against rubrics, lists of criteria graded by an LLM judge, has become a standard way to post-train language models on tasks wit

model-releasesarxiv-cs-ai
13 Aug 2026
Agents

Synchronizing Beliefs with Second-Order Theory-of-Mind in Human-Autonomy Teams (Extended Version)

DGX agent

arXiv:2608.11229v1 Announce Type: new Abstract: Comparative feedback, asking people which of two behaviors they prefer, has become a standard way to align robot and agent behavior with human intent wh

agentsarxiv-cs-ai
13 Aug 2026
Model Releases

Towards Query-Agnostic RAG Evaluation via Query Coverage and Claim Verifiability

DGX agent

arXiv:2608.11238v1 Announce Type: new Abstract: Retrieval-augmented generation improves the factuality of large language models by grounding responses in retrieved evidence, yet existing evaluation fr

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

TRACE Bench: Task-driven Roleplay Agentic Checklist Evaluation

DGX agent

arXiv:2608.11236v1 Announce Type: cross Abstract: Roleplay evaluation should do more than assign a single score: it should reveal which role requirements were tested, which failed, and which dialogue

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Transit Destination Inference from Tap-In-Only Bus Smart-Card Data: A Hierarchical Bayesian Approach

DGX agent

arXiv:2608.11223v1 Announce Type: cross Abstract: Entry-only automatic fare collection systems record boardings but not alightings, preventing direct construction of origin-destination (OD) matrices.

model-releasesarxiv-cs-lg
13 Aug 2026
Model Releases

VAKRA: Evaluating Multi-Hop Reasoning Across APIs and Retrieval Under Tool-Use Policies

DGX agent

arXiv:2608.12282v1 Announce Type: new Abstract: Agents deployed in enterprise settings must reason across structured APIs and document collections, yet existing benchmarks evaluate these capabilities

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

WavePhaseNet: A DFT-Based Method for Constructing Semantic Conceptual Hierarchy Structures (SCHS)

DGX agent

arXiv:2602.14419v2 Announce Type: cross Abstract: This paper reformulates Transformer/Attention mechanisms in Large Language Models (LLMs) through measure theory and frequency analysis, theoretically

model-releasesarxiv-cs-lg
13 Aug 2026
Model Releases

We’re launching DeepSeek-V4-Pro today! 🚀 🔷 Major Agent upgrades with strong production gains! 🔷 Flexible reasoning effort for V4-Pro & V4…

DGX agent

We’re launching DeepSeek-V4-Pro today! 🚀 🔷 Major Agent upgrades with strong production gains! 🔷 Flexible reasoning effort for V4-Pro & V4-Flash: low for simple tasks, high for daily Agent workflows, m

model-releasesdeepseek--x
13 Aug 2026
Model Releases

When Self-Consistency Backfires: Majority Vote Hurts the Majority of Hard Science Problems for Small LLMs

DGX agent

arXiv:2608.11403v1 Announce Type: new Abstract: Self-consistency (SC) via majority vote is a widely used way to spend inference-time compute: sample N chains of thought, return the plurality answer. O

model-releasesarxiv-cs-ai
13 Aug 2026
Safety

Who Would You Vote For? Auditing Political Alignment in LLMs: An Italian Case-Study

DGX agent

arXiv:2608.11649v1 Announce Type: new Abstract: As users increasingly turn to Large Language Models (LLMs) for information and advice on political matters, particularly during election periods, the po

safetyarxiv-cs-cl
13 Aug 2026
Agents

A Gateway Architecture for Enterprise MCP Authentication: Unifying Heterogeneous Auth, Identity Delegation, and the User / Non-User Persona Problem

DGX agent

arXiv:2608.10760v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) has become the de-facto interface for connecting LLM agents to enterprise tools, and adoption has been explosive: wit

agentsarxiv-cs-ai
12 Aug 2026
Model Releases

A HamNoSys-Guided Dataset and Baselines for Fine-Grained Isolated Handshape Recognition in Sign Language

DGX agent

arXiv:2608.10588v1 Announce Type: cross Abstract: Purpose: Fine-grained handshape recognition supports computational sign-language transcription, recognition, and translation, but broad, phonetically

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

A Quantum Roadmap for Softmax Attention: Exact Born-Rule Analogs for Softmax Attention on the Probability Simplex

DGX agent

arXiv:2608.11173v1 Announce Type: cross Abstract: The attention mechanism forms the foundation of many modern AI models such as the Transformer. In one subclass of problems where attention is used, in

model-releasesarxiv-cs-lg
12 Aug 2026
Model Releases

Astrolabe: Balancing Load in LLM Serving with Randomized Prediction-Guided Scheduling

DGX agent

arXiv:2508.03611v3 Announce Type: replace-cross Abstract: This paper presents Astrolabe, a randomized prediction-guided scheduler for one-shot request dispatch in multi-instance large language model (

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

b10369

DGX agent

mtmd: support pocket-tts (#26871) adapt the api text model ok working impl, need verify and clean up mtmd: build the pocket-tts transposed convolutions as GEMM + col2im ggml_conv_transpose_1d has no g

model-releasesllama-cpp-releases
12 Aug 2026
Model Releases

b10375

DGX agent

chat : tighten bare function parsing for Qwen models (#26793) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64)

model-releasesllama-cpp-releases
12 Aug 2026
Model Releases

b10398: common: add system-level config file (#26118)

DGX agent

common: Add CLI > ENV > models-presets > INI precedence CLI flags have the highest precedence ENV vars have the second-highest precedence System and User configs have the lowest precedence Linux/BSD/M

model-releasesllama-cpp-releases
12 Aug 2026
← Previous
1…494495496497498…1371
Next →