AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,661 results
24 Jul 2026

Interpretable Embeddings with Sparse Autoencoders: A Data Analysis Toolkit

ResearchDGX agent

arXiv:2512.10092v2 Announce Type: replace Abstract: Analyzing large-scale text corpora is a core challenge in machine learning, crucial for tasks like identifying undesirable model behaviors or biases

Introducing Agnost AI (YC S26): Product Analytics for AI Agents. Analytics was never built for chat & voice. And that is the new interface. …

AgentsDGX agent

Introducing Agnost AI (YC S26): Product Analytics for AI Agents. Analytics was never built for chat & voice. And that is the new interface. Your users' hidden feature requests, bugs & churn hide in co

Introducing Claude Opus 5

Model ReleasesDGX agent

Introducing Claude Opus 5 I've been offline kayaking with sea otters for much of today so I haven't had a chance to put Anthropic's new model Claude Opus 5 through its paces yet. The buzz is positive,

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Introducing Claude Opus 5 on AWS: Anthropic’s most capable Opus model

Model ReleasesDGX agent

This post covers Opus 5’s improvements and practical guidance for AI engineers integrating the model into agentic systems and production inference workloads on Amazon Bedrock. See the documentation fo

Is corruption the lobbying against Open weights?

Local AiDGX agent

Like, reading things like Anthropic 'donated' to some people with the condition of lobbying against Chinese LLMs.. it's that right? It feels nothing like freedom but at the same time it's said 'out lo

Is Deep Research Reliable? Misleading Knowledge Induces False Conclusions

Model ReleasesDGX agent

arXiv:2607.20891v1 Announce Type: new Abstract: Deep Research agents extend LLM-based assistants into long-horizon workflows involving planning, retrieval, evidence synthesis, and report generation, y

Is everyone training a single model together, based on the principle of the Tor network?

Local AiDGX agent

I just had a thought while scrolling. No idea if this already exists. What if users trained an AI model together—a bit like the Tor network or Bitcoin mining back in the day? - Participants download a

Is MoE Routing a Huffman Code? Discovering the Frequency-Diversity Law in Chain-of-Thought

Model ReleasesDGX agent

arXiv:2607.20427v1 Announce Type: cross Abstract: Mixture-of-Experts architectures have revolutionized scaling, yet the underlying logic of their routing remains a black box. In this paper, we uncover

Is Your Safe Controller Actually Safe? A Critical Review of CBF Tautologies and Hidden Assumptions

SafetyDGX agent

arXiv:2603.06954v2 Announce Type: replace Abstract: This tutorial provides a critical review of the practical application of Control Barrier Functions (CBFs) in robotic safety. While the theoretical f

Isolating LLM Alignment from Regex: Zero Coverage and Metric-Dependent Divergence Under Adversarial Mutation

Model ReleasesDGX agent

arXiv:2607.20494v1 Announce Type: new Abstract: Production LLM applications commonly stack a regex filter in front of model-side alignment; prior work found no measurable coverage gain from adding a l

IssueTrojanBench: Benchmarking AI Coding Agents Against Malicious Issue Requests

Model ReleasesDGX agent

arXiv:2607.20759v1 Announce Type: cross Abstract: AI coding agents powered by LLMs are increasingly integrated into real-world software development, where they generate, edit, and execute code with au

It appears that the anti opensource AI lobby is far outgunned already

Local AiDGX agent

The earlier post on this subreddit by 20+ companies signing the petition including Microsoft, Meta, Nvidia, YC (https://www.microsoft.com/en-us/corporate-responsibility/topics/open-weight/) etc plus t

JAXBench: Benchmarking Autonomous TPU Kernel Optimization

Model ReleasesDGX agent

arXiv:2607.20466v1 Announce Type: new Abstract: Rigorous benchmarks have driven progress in autonomous GPU kernel performance optimization by establishing a shared target to hillclimb on, but no equiv

Joint Utilization of Geospatial and census proxies for Autoencoder-Assisted Downscaling (JUGAAD) of socioeconomic indicators in India

ApplicationsDGX agent

arXiv:2607.20559v1 Announce Type: cross Abstract: Monitoring poverty and food security indicators is imperative for addressing socioeconomic challenges in developing nations. A limitation is mismatche

Just over 6 months later, Opus 5 now produces near-superhuman level spreadsheets and slide decks that match what a consultant would make. Th…

Model ReleasesDGX agent

Just over 6 months later, Opus 5 now produces near-superhuman level spreadsheets and slide decks that match what a consultant would make. Things are changing fast. Media I'm hearing from many folks ac

KeySI: An Interaction Framework for Tuning Text Embeddings Based on Human Feedback

SafetyDGX agent

arXiv:2607.20556v1 Announce Type: new Abstract: In large-scale text analysis tasks, pre-trained language models are often used to embed text corpora for downstream analysis. However, such models may s

.@Kimi_Moonshot K3 lands on Together on Monday! We ran 452 DeepSWE rollouts against Claude Fable 5: near-flagship coding at ~35% of the pric…

Model ReleasesDGX agent

.@Kimi_Moonshot K3 lands on Together on Monday! We ran 452 DeepSWE rollouts against Claude Fable 5: near-flagship coding at ~35% of the price, and K3 pulls ahead at higher pass@k's. Full deep-dive: ht

Knowledge Injection Exists in MoE? Exploring Expert-Aware Contrast Decoding in MoE for Mitigating LLMs'Hallucinations

ResearchDGX agent

arXiv:2607.20426v1 Announce Type: cross Abstract: Existing LLM hallucination mitigation methods, including prompt engineering and model optimization, either hardly alter models'internal knowledge or h

KroQuant: Kronecker-Structured Block Transforms for Efficient Post-Training Quantization of Diffusion Transformers

HardwareDGX agent

arXiv:2607.21446v1 Announce Type: cross Abstract: Post-training quantization (PTQ) of diffusion transformers (DiTs) to W4A4 severely degrades output quality, because activations entering each linear l

Latent Variable-Mediated Cross-Learning for Few-Shot Acoustic Impedance Imaging

ResearchDGX agent

arXiv:2607.20989v1 Announce Type: new Abstract: Acoustic impedance imaging is a fundamental yet severely ill-posed problem in subsurface analysis: the seismic wavelet is unknown, observations are band

Lawmakers call for a ‘kill switch’ after rogue AI causes alarm

IndustryDGX agent

U.S. lawmakers are preparing to introduce a bill that will give the government the ability to trigger an emergency shutdown of artificial intelligence models that may cause harm to the public. In what

LEAD: Breaking the No-Recovery Bottleneck in Long-Horizon Reasoning

ResearchDGX agent

Long-horizon execution in Large Language Models (LLMs) remains unstable even when high-level strategies are provided. Evaluating on controlled algorithmic puzzles, we demonstrate that while decomposit

Leaky Language Models: Stealing Architecture and Inference Optimizations via Per-Token Timing

Model ReleasesDGX agent

arXiv:2607.20723v1 Announce Type: cross Abstract: This work presents LeakyLMs, a set of attacks that leak proprietary model, architecture, and deployment information from production language models. L

LeanFlow: A Case Study in Workflow-Driven Lean Autoformalization

AgentsDGX agent

arXiv:2607.20503v1 Announce Type: new Abstract: We present and evaluate LeanFlow, an LLM agent system specialized for translating mathematical papers into buildable Lean projects. Recent verifier-in-t

Learn2Zinc: Fine-tuning Small Language Models for Text-to-Model Translation in MiniZinc

Model ReleasesDGX agent

arXiv:2607.20456v1 Announce Type: cross Abstract: Large language models excel at code generation for mainstream programming languages but struggle with rare, domain-specific languages such as MiniZinc

Learning-based Seam Correspondence Reconstruction in Sewing Patterns

ResearchDGX agent

arXiv:2607.21213v1 Announce Type: new Abstract: Digital sewing patterns typically consist of disjoint 2D panels without explicit stitch annotations, making downstream 3D modeling reliant on labor-inte

Learning causality from internet videos in latent space first, and then using RL to teach the foundation model how to act. This approach is …

Model ReleasesDGX agent

Learning causality from internet videos in latent space first, and then using RL to teach the foundation model how to act. This approach is 30× cheaper than Gemini 3.1 Flash on pretraining and achieve

Learning to Detect UI Principle Violations via Reinforcement Learning

ApplicationsDGX agent

arXiv:2607.20690v1 Announce Type: new Abstract: Small language models and coding agents increasingly generate web front-end code, yet their outputs are typically evaluated primarily for functional cor

Learning to Navigate Efficiently with Only 0.58M Trainable Parameters

Model ReleasesDGX agent

arXiv:2607.11029v2 Announce Type: replace-cross Abstract: Recent progress in visual navigation has largely been driven by scale: end-to-end policies with hundreds of millions of parameters trained on

LegalCiteTrust: Benchmarking Citation Trustworthiness in Chinese Long-Form Legal Research Reports

Model ReleasesDGX agent

arXiv:2607.20872v1 Announce Type: new Abstract: Long-form legal research reports increasingly rely on LLMs and agentic research systems, but their reliability depends not only on answering the task, b

Lessons and Open Questions from a Unified Study of Camera-Trap Species Recognition Over Time

Model ReleasesDGX agent

arXiv:2603.20509v2 Announce Type: replace Abstract: Camera traps are vital for large-scale biodiversity monitoring, yet accurate automated analysis remains challenging due to diverse deployment enviro

Let’s build seven trillion dollars worth of data centers for sports betting!

SafetyDGX agent

Let’s build seven trillion dollars worth of data centers for sports betting! Statistics of week: “More Americans pay for sports betting apps (5%) than pay for AI… 37% of consumers say none of AI's use

Leveraging Biokinetic Knowledge Priors for Data-Scarce Bioprocess Modeling

TutorialsDGX agent

arXiv:2607.20539v1 Announce Type: cross Abstract: While deep learning has accelerated drug discovery, its impact on biomanufacturing has been considerably more limited. The reason is data scarcity. Bi

LinearARD: Linear-Memory Attention Distillation for RoPE Restoration

ResearchDGX agent

arXiv:2604.00004v2 Announce Type: replace-cross Abstract: The extension of context windows in Large Language Models is typically facilitated by scaling positional encodings followed by lightweight Con

Live now: our entire AI x Evals Track! https://www.youtube.com/watch?v=q2JrUKBMf0w&list=PLJ7eF79yCUHc - @aparnadhinak, CPO @arizeai - Jason …

ToolsDGX agent

Live now: our entire AI x Evals Track! https://www.youtube.com/watch?v=q2JrUKBMf0w&list=PLJ7eF79yCUHc - @aparnadhinak, CPO @arizeai - Jason Lopatecki, CEO @arizeai - @lukaspet, Co-Founder, Andon Labs

LLM-INSTRUCT at UZH Shared Task 2026: Constraint-Aware Retrieval and Selective Debate for Paragraph-Level Argument Mining

AgentsDGX agent

arXiv:2607.20430v1 Announce Type: cross Abstract: We present LLM-INSTRUCT, the winning system for the UZH Shared Task at ArgMining 2026 on paragraph-level argument mining in UN and UNESCO resolutions.

LLMs Get Lost in Evolving User Intent

TutorialsDGX agent

arXiv:2607.20734v1 Announce Type: new Abstract: As LLMs become more capable, they are increasingly deployed as collaborative agents, taking on user-delegated tasks through iterative interaction. Yet g

Logic Programming Semantics for Causal Processes

ResearchDGX agent

arXiv:2607.21233v1 Announce Type: new Abstract: Motivated by challenging modelling issues in the life sciences, we investigate the relationship between logic programming semantics and the eventual sta

Logical Regression for Planning with Axioms

ResearchDGX agent

arXiv:2607.21414v1 Announce Type: new Abstract: In automated planning, logical regression is an operation that returns the most general condition necessary for an action to achieve a particular formul

Loss-Complexity Landscape and Model Structure Functions

ResearchDGX agent

arXiv:2507.13543v5 Announce Type: replace-cross Abstract: We develop a framework for dualizing the Kolmogorov structure function h_x(alpha), which then allows using computable complexity proxies. We e

Loss Landscape Topology Reveals Why Simple Baselines are Competitive at 3D Point Cloud Segmentation Under Class Imbalance

ResearchDGX agent

arXiv:2607.21089v1 Announce Type: new Abstract: Semantic segmentation of 3D point clouds faces severe class imbalance, yet the effectiveness of specialized imbalance-aware methods from 2D computer vis

M^3-Gen: Interpretable Multimodal Generation of Gene Expression Profiles Using Clinical and Imaging Data

TutorialsDGX agent

arXiv:2607.21343v1 Announce Type: cross Abstract: Integrating heterogeneous biomedical data, including clinical metadata, histopathology images, and molecular profiles, is crucial for comprehensive di

Machine-Learned Compact Subspace Generation for Quantum Selected Configuration Interaction within Density Matrix Embedding Framework

TutorialsDGX agent

arXiv:2607.20585v1 Announce Type: cross Abstract: Sample-based Quantum Diagonalization (SQD), an extension of Quantum Selected Configuration Interaction (QSCI), has emerged as a promising hybrid quant

Machine Learning for Charge State Characterization of Isolated Double Quantum Dots

Local AiDGX agent

arXiv:2607.20871v1 Announce Type: cross Abstract: Scaling semiconductor quantum dot arrays toward fault-tolerant quantum computing requires efficient tuneup of spin qubits, a process that depends on t

MAGE-Vein: Multi-Instance Age and Gender Estimation from Finger Vein Images

ResearchDGX agent

arXiv:2607.20897v1 Announce Type: new Abstract: Age estimation from finger vein images has been widely considered impractical due to severe demographic biases in public datasets and physiological conf

MagicMakeup: A Region-Controllable Diffusion Transformer for High-Fidelity Makeup-Transfer

Model ReleasesDGX agent

arXiv:2607.20924v1 Announce Type: new Abstract: Makeup-transfer applies the reference makeup to the source face while preserving the source identity. Despite advances in full-face editing by diffusion

Making Open-Source Text LLM Watermarks Durable Against Merging

TutorialsDGX agent

arXiv:2607.20435v1 Announce Type: cross Abstract: Open-source LLMs (OSMs)arereaching near state-of-the-art performance, prompting prior works to trace the text they generate by embedding text watermar

Marking the Wrong Symptoms: Evaluating LLM Watermarks in Medical Texts

ResearchDGX agent

arXiv:2607.20462v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into clinical workflows, stressing the need for reliable traceability of model-generated output

Masked Topology Modeling for Self-Supervised Learning on Parametric CAD

ResearchDGX agent

arXiv:2607.20642v1 Announce Type: new Abstract: Computer aided design (CAD) is ubiquitous: virtually any modern object was designed using editable CAD tools. However, with the shortage of available CA

Mean-to-Score Discrete Diffusion: Posterior-Mean Denoisers for Score Entropy

Model ReleasesDGX agent

arXiv:2607.21372v1 Announce Type: cross Abstract: Score Entropy Discrete Diffusion (SEDD) parameterizes discrete reverse processes with unconstrained positive score ratios. While positivity guarantees

MedGame: Storytelling Gamification Empowered by Large Language Models for Medical Education

Model ReleasesDGX agent

arXiv:2607.21570v1 Announce Type: new Abstract: Large Language Models (LLMs) show promise for medical education, but most existing systems focus on localized interactions such as question answering or

MELLA: Bridging Linguistic Capability and Cultural Groundedness for Low-Resource Language MLLMs

SafetyDGX agent

arXiv:2508.05502v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) perform strongly in high-resource languages, yet often produce fluent but culturally 'thin' descripti

Memoir: Should a Model Write to Its Memory While It Thinks?

ResearchDGX agent

arXiv:2607.20792v1 Announce Type: new Abstract: Memoir combines per-sample fast memory, shared slow parameters, variable-depth latent recurrence, and a future-latent energy objective. We test its risk

Memory-Computation Tradeoffs in Semi Amortized Parametric Optimization

Model ReleasesDGX agent

arXiv:2607.20769v1 Announce Type: new Abstract: Learning-enabled decision systems often use offline data or computation to reduce online compute cost. Despite the empirical success of such approaches,

MemTools: A Unified Research Framework for Interoperable Agent Memory

Model ReleasesDGX agent

arXiv:2607.21404v1 Announce Type: new Abstract: While memory systems are essential for agent architectures, pervasive architectural fragmentation restricts systematic research. Existing implementation

Meta is making its AI chatbot more like an assistant

Model ReleasesDGX agent

Meta is upgrading its AI chatbot with new productivity features in a bid to compete with rivals like Gemini, ChatGPT, and Claude. The update will allow Meta AI to tap into your calendar to help you pl

Meta makes Muse Spark 1.1 available to consumers, debuts new Facebook features

IndustryDGX agent

Meta Platforms Inc. today made its latest large language model available to consumers through its Meta AI chatbot. The update is rolling out alongside several enhancements to Facebook. Meta’s flagship

Meta, Nvidia, Microsoft, a16z, and others sign a letter defending open-source AI; Jensen Huang, in his first X post, says open models strengthen cybersecurity (Leo Schwartz/The Information)

HardwareDGX agent

Leo Schwartz / The Information: Meta, Nvidia, Microsoft, a16z, and others sign a letter defending open-source AI; Jensen Huang, in his first X post, says open models strengthen cybersecurity — Many of

Meta updates Meta AI with Muse Spark 1.1-powered agentic capabilities, connecting to Gmail and Google Calendar to perform tasks like creating daily updates (Ina Fried/Axios)

AgentsDGX agent

Ina Fried / Axios: Meta updates Meta AI with Muse Spark 1.1-powered agentic capabilities, connecting to Gmail and Google Calendar to perform tasks like creating daily updates — Meta is giving its AI a

MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference

AgentsDGX agent

arXiv:2607.20507v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for program-aided reasoning, agentic decision making, and structured task execution, but these applic

← Previous
1…210211212213214…1412
Next →