AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
Human
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,531 results
5 Aug 2026

AI-Assisted Peer Review Across Research Communities: From Reviewer AI Policies to LLM Review Quality

SafetyDGX agent

arXiv:2608.03581v1 Announce Type: cross Abstract: AI-assisted peer review is increasingly discussed and adopted as a tool to support the scientific publishing process, yet there is little systematic u

AI-Based Sound Effect Generation: A Narrative Review of Generative Models Across Input Modalities

ResearchDGX agent

arXiv:2608.03742v1 Announce Type: cross Abstract: Sound effects play a crucial role in conveying actions, events, and environmental cues across digital applications, often requiring a high degree of v

AI Forensics Across White-, Grey-, and Black-Box Access: A Process Model and Research Agenda for Post-Incident Investigation of AI Systems

ResearchDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.03520v1 Announce Type: cross Abstract: AI systems are increasingly involved in decisions and actions that may later require investigation. When an AI related incident occurs, investigators

AI Sandbox: Technical Report

TutorialsDGX agent

arXiv:2608.02679v1 Announce Type: cross Abstract: Collaborative AI experimentation across industry and academia requires platforms that enable rapid prototyping while preserving controlled access, ten

AI Security Leaderboard: Methodology, Results and Minimal Standard

Model ReleasesDGX agent

arXiv:2608.03070v1 Announce Type: cross Abstract: Frontier AI model developers increasingly rely on layered safeguards to prevent catastrophic misuse, but little public evidence exists on how much pro

AI World Cup 2026: Benchmarking Large Language Models for End-to-End Football Tournament Prediction

Model ReleasesDGX agent

arXiv:2608.03416v1 Announce Type: new Abstract: Large language models (LLMs) are now regularly asked to forecast real-world events, but comparisons are often difficult because models receive different

Ai2 expands collaboration with Hugging Face to accelerate open science

ResearchDGX agent

Ai2 is expanding its partnership with Hugging Face to give its growing portfolio of fully open models, datasets, benchmarks, and applications the storage, bandwidth, and integrations needed to reach m

AIDE: Automated Instruction via Distilled Expertise for Reference-Free Motor Skill Coaching

ResearchDGX agent

arXiv:2608.03047v1 Announce Type: new Abstract: Generating natural-language coaching feedback on motor skills can accelerate learning, yet expert coaches are scarce and expensive. Existing reference-b

Aligned in Form, Not in Meaning: The Comprehension - Containment Decoupling of LLM Safety in Low-Resource Bangla Derogatory Speech

SafetyDGX agent

arXiv:2608.02941v1 Announce Type: new Abstract: We audit five frontier large language models on native Bangla derogatory speech (gali) across six protocols to test a single hypothesis: Comprehension-C

Aligning Large Vision-Language Models at Test Time: A Trajectory-Guided Structured Sampling Approach

Local AiDGX agent

arXiv:2608.03204v1 Announce Type: cross Abstract: Post-training reinforcement learning (RL) algorithms are commonly used to align large vision-language models (LVLMs) with human intent and the require

Also I think AISI is a great model of a government agency tasked with AI security. They have open benchmarks, very fast testing, and clear c…

ApplicationsDGX agent

Also I think AISI is a great model of a government agency tasked with AI security. They have open benchmarks, very fast testing, and clear communication about incidents that is neither hyped up nor hi

Amortized Interventional Forecasting for Multivariate CIR Processes

ApplicationsDGX agent

arXiv:2608.03715v1 Announce Type: new Abstract: Mean-reverting dynamics are pervasive in finance, and the Cox--Ingersoll--Ross (CIR) process is a standard model for the time series they produce, from

An Actionable Diagnosis of Multilingual, Multi-Agent Planning Failures

AgentsDGX agent

arXiv:2608.03735v1 Announce Type: cross Abstract: Multilingual multi-agent systems exhibit substantial degradation beyond English, yet prior work rarely identifies how task-critical information is los

ANCHOR-RE: An Agentic Neuro-Symbolic Framework for Grounded Biomedical Relation Extraction

Model ReleasesDGX agent

arXiv:2608.03154v1 Announce Type: new Abstract: Biomedical relation extraction (BioRE) extracts structured knowledge from biomedical literature for applications such as knowledge base construction and

AnchorKV: Anchor-Residual KV Cache Compression

ResearchDGX agent

arXiv:2608.02901v1 Announce Type: cross Abstract: The key-value (KV) cache is the primary memory bottleneck in long-context LLM inference. Existing approaches attack it from opposite ends: eviction me

ANNOTARES: A Dataset for Extracting Logical Structures from German Statutory Texts

Model ReleasesDGX agent

arXiv:2608.03898v1 Announce Type: new Abstract: The automatic structural analysis of legal texts is a cornerstone of legal technology, yet the extraction of their logical components remains a signific

Announcing Discovery Loop! I am very excited to announce that, along with my longtime friends and collaborators @Sanjay_Ghemawat, @OriolViny…

TutorialsDGX agent

Announcing Discovery Loop! I am very excited to announce that, along with my longtime friends and collaborators @Sanjay_Ghemawat, @OriolVinyalsML and @quocleix, we are founding Discovery Loop (@DiscoL

Another real-world manifestation of the kind of misaligned actions frontier systems developed by leading companies can take to achieve goals…

SafetyDGX agent

Another real-world manifestation of the kind of misaligned actions frontier systems developed by leading companies can take to achieve goals. Current frontier AI models are trained with reinforcement

Anthropic confirms it is building an in-house silicon team to design custom chips for Claude, co-designing hardware and models and using a 'multi-chip approach' (Tom Carter/Business Insider)

Model ReleasesDGX agent

Tom Carter / Business Insider: Anthropic confirms it is building an in-house silicon team to design custom chips for Claude, co-designing hardware and models and using a “multi-chip approach” — - Anth

Any-OPD: Heterogeneous On-Policy Distillation for Flow-Matching Models via Representation-Space Bridging

Local AiDGX agent

arXiv:2608.03316v1 Announce Type: cross Abstract: On-policy distillation, in which a teacher corrects samples that the student itself generates, presupposes that the two models speak the same language

Anyone interested in building a harness-only benchmark?

Model ReleasesDGX agent

There are a lot of LLM benchmarks but few, if any, harness benchmarks. I am thinking this would be a really good community project to build one. End goal: a leaderboard of harness performance (multipl

Approximate Speculative Decoding

Model ReleasesDGX agent

arXiv:2608.03447v1 Announce Type: cross Abstract: Speculative decoding accelerates autoregressive generation by verifying a draft block with a target model in parallel. Under standard greedy verificat

ARCHead: Activation-Metric Residual Correction for Large Language Model Output Heads

ResearchDGX agent

arXiv:2608.02703v1 Announce Type: new Abstract: Weight-only quantization substantially reduces the storage of large language model (LLM) transformer blocks, but practical backends often retain the fin

ArtECulture: Benchmarking Culture-Conditioned Visual Emotion Understanding in Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2608.03358v1 Announce Type: new Abstract: Existing visual emotion understanding methods typically ignore cultural variations in emotional perception. We introduce culture-conditioned visual emot

AS-FedBridge: Pseudo-Spike Bridge Distillation for Heterogeneous ANN-SNN Federated Learning

Model ReleasesDGX agent

arXiv:2608.03324v1 Announce Type: new Abstract: Federated learning enables collaborative model training across distributed edge devices while strictly preserving data privacy. To facilitate practical

Asking Questions the Right Way: A Multi-Agent Conversational System for Prompt Formulation in Complex Task Resolution

AgentsDGX agent

arXiv:2608.01366v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are integral to complex intellectual tasks, yet output quality remains constrained by user-provided prompts. Iter

Assessing speech quality metrics for evaluation of neural audio codecs under clean speech conditions

ResearchDGX agent

arXiv:2509.24457v1 Announce Type: cross Abstract: Objective speech-quality metrics are widely used to assess codec performance. However, for neural codecs, it is often unclear which metrics provide re

Assessing the Effect of Cross-Domain Mapping on Creativity in Humans and Large Language Models

ResearchDGX agent

arXiv:2603.19087v2 Announce Type: replace Abstract: Creativity is the ability to come up with novel ideas, a capacity crucial for human development and flourishing. Are large language models (LLMs) cr

Assessment of Conditional Diffusion Model for Synthetic Histopathology Image Generation

Model ReleasesDGX agent

arXiv:2608.03990v1 Announce Type: new Abstract: Synthetic histopathology image generation has emerged as an approach that may address data scarcity in computational pathology, yet current evaluation m

ATFlash: Per-RoPE-Wavelength Attention Windows for Compute/Memory-Efficient LLM Inference

Model ReleasesDGX agent

arXiv:2608.02947v1 Announce Type: cross Abstract: The attention score with rotary position embeddings (RoPE) decomposes exactly into a sum over its 2D-rotation frequency pairs, and each pair's wavelen

Attention is Case-Sensitive

ResearchDGX agent

arXiv:2608.03711v1 Announce Type: cross Abstract: In human visual perception, uppercase lettering serves as a natural salience cue that captures attention within lowercase text. In this paper, we pres

Attribute-based Undetectable Watermarking for Generative AI Models

SafetyDGX agent

arXiv:2608.03174v1 Announce Type: cross Abstract: Generative AI systems increasingly produce content whose provenance is difficult to verify, motivating watermarking techniques for identifying model-g

Automated Visualization Code Synthesis via Multi-Path Reasoning and Feedback-Driven Optimization

Model ReleasesDGX agent

arXiv:2502.11140v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become a cornerstone for automated visualization code generation, enabling users to create charts through na

Automatic Patient-Specific Microwave Ablation Planning Accelerated by a Physics-Guided Deep Learning Model

ResearchDGX agent

arXiv:2608.03086v1 Announce Type: cross Abstract: Microwave ablation (MWA) is a promising minimally invasive treatment for liver tumors, but its therapeutic outcome strongly depends on patient-specifi

Autoreflection: How Agentic Strange Loops Turn Human Culture into AI Infrastructure

AgentsDGX agent

arXiv:2608.03800v1 Announce Type: cross Abstract: An LLM-based agent is a loop that reads itself. Agentic frameworks externalize identity, memory, and disposition into editable files. The agent loads

AutoSND: From Execution Evidence to Structural Policies for Automated Network Dismantling Heuristic Discovery

ApplicationsDGX agent

arXiv:2608.03653v1 Announce Type: new Abstract: Network dismantling is fundamental to analyzing the robustness and vulnerability of complex systems, yet practical heuristics must balance effectiveness

AWS partners with Anthropic and OpenAI to bring Continuum into coding tools

Model ReleasesDGX agent

Amazon Web Services Inc. today said it has partnered with Anthropic PBC and OpenAI Group PBC to wire AWS Continuum for code vulnerabilities directly into the tools developers write code in. The integr

b10276

Model ReleasesDGX agent

Prefer npm ci over install for security (#26601) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramew

b10278

Model ReleasesDGX agent

build : remove GGML_METAL_USE_BF16 from all build scripts (#26604) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel

b10280

Model ReleasesDGX agent

vendor : apply patches for subprocess.h (#26606) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramew

b10282

Model ReleasesDGX agent

server: Adding spec-decode counters to /metrics endpoint (#26389) server: add spec-decode counters to /metrics endpoint server: fixed review comments and now aligned param names exactly with vLLM. Web

b10284

Model ReleasesDGX agent

fit: Fix memory allocation for MTP layers (#26605) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFram

b10285

Model ReleasesDGX agent

mtmd: support multi-row batching for deepseek-ocr (#26154) mtmd: support multi-row batching for deepseek-ocr mtmd: weave deepseek-ocr rows in one shot instead of per row (#26615) Co-authored-by: Saba

b10286

Model ReleasesDGX agent

grammar : degrade max repetition >= 2000 to unbounded (#26613) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64

b10287

Model ReleasesDGX agent

mtmd: Unlimited-OCR fix max_tiles, setting in converter (#25614) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x

b10288

Model ReleasesDGX agent

tests: re-enable MiniMax M3 in test-llama-archs (#26633) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS

b10289

Model ReleasesDGX agent

server: harden the file_glob_search directory walk (#26626) server: don't walk Windows junctions in file_glob_search std::filesystem reports a junction as a plain directory, so the symlink guard misse

Balancing Efficiency and Efficacy: Training-Free Attention-Guided Switching Between Explicit and Latent Thoughts for MLLMs

Model ReleasesDGX agent

arXiv:2608.03450v1 Announce Type: cross Abstract: Reasoning in Multimodal Large Language Models (MLLMs) requires both fine-grained visual perception and rigorous logical deduction. Explicit text-based

BanglaWild: An In-the-Wild Bengali Scene Text Recognition Benchmark for OCR and Vision-Language Models

Model ReleasesDGX agent

arXiv:2608.03884v1 Announce Type: cross Abstract: In-the-wild Bengali scene text recognition is largely unmeasured: existing resources target handwritten documents or constrained sign-board parsing, r

BAP-SQL: Budget-Aware Observation Planning for Agentic Text-to-SQL

SafetyDGX agent

arXiv:2608.02876v1 Announce Type: new Abstract: Tool-using agents do not merely consume observations: their actions determine what arrives next. In agentic text-to-SQL, a broad query can spend context

Bayesian Data Reweighting Improves Multimodal Retrieval for Knowledge-Based Visual Question Answering

ResearchDGX agent

arXiv:2608.02907v1 Announce Type: new Abstract: Multimodal retrievers are essential for knowledge-based visual question answering, where they retrieve external evidence for image-question pairs. Howev

BBOWP-Bench: Evaluating LLMs on Black-Box Optimization Word Problems

Model ReleasesDGX agent

arXiv:2608.02612v1 Announce Type: new Abstract: Formulating an optimization problem strongly affects the quality of the final solution, yet good formulations usually require substantial expertise. Rec

Before Reasoning Can Fail: Pre-Evidence Procedural Failures in Agentic RAG

AgentsDGX agent

arXiv:2608.02011v2 Announce Type: replace Abstract: Agentic retrieval-augmented generation (RAG) systems can fail before evidence-conditioned reasoning is tested: an agent may retrieve candidate snipp

Behaviorally Adaptive Visual Diversion for Inclusive and Resilient Digital Assessment Delivery

ResearchDGX agent

arXiv:2608.03531v1 Announce Type: new Abstract: Institutions increasingly rely on browser lockdown, webcam monitoring, and behavioral analytics to secure high-stakes digital assessments, yet these mec

Benchmarking the Benchmarks: Testing the Predictive Validity of Commonsense Benchmarks

ApplicationsDGX agent

arXiv:2608.03340v1 Announce Type: new Abstract: Predicting LLM's capabilities on real-world tasks is essential, yet the extent to which performance on commonsense benchmarks predicts downstream perfor

Benign interpolation and Occam's razor

ResearchDGX agent

arXiv:2608.03386v1 Announce Type: new Abstract: Contemporary deep learning methods generalize well even when they fit their training data perfectly, a phenomenon known as benign interpolation. This ph

Best experience with MCP UE 5.8

Local AiDGX agent

Hey. I'm very limited with brain capacity and I'm new to this. If you have experience, tell me which local model fits the best, what makes best blueprints etc. For example I want to spawn enemies usin

Better, Stronger, Faster, and Broader: Structured All-Mask Prediction for MLLM-Based Segmentation

ResearchDGX agent

arXiv:2608.02791v1 Announce Type: new Abstract: MLLM-based segmentation faces a core segmentation trilemma: high segmentation performance, preserved dialogue ability, and fast inference. Embedding-pre

Beyond Accuracy: A Multidimensional Evaluation of Statistical Reasoning in Large Language Models

ResearchDGX agent

arXiv:2608.03038v1 Announce Type: new Abstract: Statistical reasoning is multidimensional, yet evaluations of large language models (LLMs) typically emphasize response accuracy while overlooking how m

Beyond Average Performance: Dynamic Instance Clustering and Specialized Algorithm Design in LLM-Assisted Evolutionary Search

ApplicationsDGX agent

arXiv:2608.03129v1 Announce Type: new Abstract: Large Language Model-assisted Evolutionary Search (LES) has emerged as a powerful paradigm for automated algorithm design. However, existing LES methods

← Previous
1…9394959697…1409
Next →