AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
20 Apr 2026

Calling all builders 🛠️📣 Google AI Pro and Ultra subscribers will now get increased usage limits and access to Nano Banana Pro and Gemini …

Model ReleasesDGX agent

Calling all builders 🛠️📣 Google AI Pro and Ultra subscribers will now get increased usage limits and access to Nano Banana Pro and Gemini Pro models in @GoogleAIStudio — no API key required. Sign in w

Capture the Flags: Family-Based Evaluation of Agentic LLMs via Semantics-Preserving Transformations

AgentsDGX agent

arXiv:2602.05523v2 Announce Type: replace-cross Abstract: Agentic large language models (LLMs) are increasingly evaluated on cybersecurity tasks using capture-the-flag (CTF) benchmarks, yet existing p

Consistency Analysis of Sentiment Predictions using Syntactic & Semantic Context Assessment Summarization (SSAS)

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2604.15547v1 Announce Type: cross Abstract: The fundamental challenge of using Large Language Models (LLMs) for reliable, enterprise-grade analytics, such as sentiment prediction, is the conflic

Context, from the time of the o1-preview launch: https://www.oneusefulthing.org/p/something-new-on-openais-strawberry

ApplicationsDGX agent

Ethan Mollick shared context about OpenAI's o1-preview model launch, likely discussing the capabilities and implications of this new reasoning-focused AI system that represents a shift toward models d

CRoCoDiL: Continuous and Robust Conditioned Diffusion for Language

ResearchDGX agent

arXiv:2603.20210v3 Announce Type: replace-cross Abstract: Masked Diffusion Models (MDMs) provide an efficient non-causal alternative to autoregressive generation but often struggle with token dependen

DPrivBench: Benchmarking LLMs' Reasoning for Differential Privacy

Model ReleasesDGX agent

arXiv:2604.15851v1 Announce Type: cross Abstract: Differential privacy (DP) has a wide range of applications for protecting data privacy, but designing and verifying DP algorithms requires expert-leve

Enhancing AI and Dynamical Subseasonal Forecasts with Probabilistic Bias Correction

SafetyDGX agent

arXiv:2604.16238v1 Announce Type: new Abstract: Decision-makers rely on weather forecasts to plant crops, manage wildfires, allocate water and energy, and prepare for weather extremes. Today, such for

'Excuse me, may I say something...' CoLabScience, A Proactive AI Assistant for Biomedical Discovery and LLM-Expert Collaborations

Model ReleasesDGX agent

arXiv:2604.15588v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into scientific workflows presents exciting opportunities to accelerate biomedical discovery. However,

Exploring the Capability Boundaries of LLMs in Mastering of Chinese Chouxiang Language

Model ReleasesDGX agent

arXiv:2604.15841v1 Announce Type: new Abstract: While large language models (LLMs) have achieved remarkable success in general language tasks, their performance on Chouxiang Language, a representative

Federated Learning with Quantum Enhanced LSTM for Applications in High Energy Physics

Local AiDGX agent

arXiv:2604.15775v1 Announce Type: new Abstract: Learning with large-scale datasets and information-critical applications, such as in High Energy Physics (HEP), demands highly complex, large-scale mode

HiPreNets: High-Precision Neural Networks through Progressive Training

Model ReleasesDGX agent

arXiv:2506.15064v3 Announce Type: replace Abstract: Deep neural networks are powerful tools for solving nonlinear problems in science and engineering, but training highly accurate models becomes chall

HyCal: A Training-Free Prototype Calibration Method for Cross-Discipline Few-Shot Class-Incremental Learning

Model ReleasesDGX agent

arXiv:2604.15678v1 Announce Type: new Abstract: Pretrained Vision-Language Models (VLMs) like CLIP show promise in continual learning, but existing Few-Shot Class-Incremental Learning (FSCIL) methods

Is this chart lying to me? Automating the detection of misleading visualizations

Model ReleasesDGX agent

arXiv:2508.21675v3 Announce Type: replace Abstract: Misleading visualizations are a potent driver of misinformation on social media and the web. By violating chart design principles, they distort data

JFinTEB: Japanese Financial Text Embedding Benchmark

Model ReleasesDGX agent

arXiv:2604.15882v1 Announce Type: cross Abstract: We introduce JFinTEB, the first comprehensive benchmark specifically designed for evaluating Japanese financial text embeddings. Existing embedding be

Kimi K2.6 just dropped. And it crushed Claude Opus 4.6 on SWE-Bench Pro. Kimi K2.6: 58.6 GPT-5.4 xhigh: 57.7 Gemini 3.1 Pro: 54.2 Claude Opu…

Model ReleasesDGX agent

Kimi K2.6 just dropped. And it crushed Claude Opus 4.6 on SWE-Bench Pro. Kimi K2.6: 58.6 GPT-5.4 xhigh: 57.7 Gemini 3.1 Pro: 54.2 Claude Opus 4.6: 53.4 An open source Chinese model is now #1 on agenti

LaMSUM: Amplifying Voices Against Harassment through LLM Guided Extractive Summarization of User Incident Reports

Model ReleasesDGX agent

arXiv:2406.15809v5 Announce Type: replace Abstract: Citizen reporting platforms help the public and authorities stay informed about sexual harassment incidents. However, the high volume of data shared

llm-openrouter 0.6

ToolsDGX agent

Release: llm-openrouter 0.6 llm openrouter refresh command for refreshing the list of available models without waiting for the cache to expire. I added this feature so I could try Kimi 2.6 on OpenRout

Mamba-SSM with LLM Reasoning for Feature Selection: Faithfulness-Aware Biomarker Discovery

Model ReleasesDGX agent

arXiv:2604.14334v2 Announce Type: replace-cross Abstract: Gradient saliency from deep sequence models surfaces candidate biomarkers efficiently, but the resulting gene lists can be contaminated by tis

Mapping High-Performance Regions in Battery Scheduling across Data Uncertainty, Battery Design, and Planning Horizons

ResearchDGX agent

arXiv:2604.15360v1 Announce Type: new Abstract: This study presents a triadic analysis of energy storage operation under multi-stage model predictive control, investigating the interplay between data

Mind DeepResearch Technical Report

Model ReleasesDGX agent

arXiv:2604.14518v2 Announce Type: replace Abstract: We present Mind DeepResearch (MindDR), an efficient multi-agent deep research framework that achieves leading performance with only ~30B-parameter m

Mind's Eye: A Benchmark of Visual Abstraction, Transformation and Composition for Multimodal LLMs

Model ReleasesDGX agent

arXiv:2604.16054v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have achieved impressive progress on vision language benchmarks, yet their capacity for visual cognitive and

NeuroMesh: A Unified Neural Inference Framework for Decentralized Multi-Robot Collaboration

HardwareDGX agent

arXiv:2604.15475v1 Announce Type: new Abstract: Deploying learned multi-robot models on heterogeneous robots remains challenging due to hardware heterogeneity, communication constraints, and the lack

Optimizing Korean-Centric LLMs via Token Pruning

Model ReleasesDGX agent

arXiv:2604.16235v1 Announce Type: new Abstract: This paper presents a systematic benchmark of state-of-the-art multilingual large language models (LLMs) adapted via token pruning - a compression techn

Polarization by Default: Auditing Recommendation Bias in LLM-Based Content Curation

Model ReleasesDGX agent

arXiv:2604.15937v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed to curate and rank human-created content, yet the nature and structure of their biases in these

Power to the Clients: Federated Learning in a Dictatorship Setting

Local AiDGX agent

arXiv:2510.22149v3 Announce Type: replace-cross Abstract: Federated learning (FL) has emerged as a promising paradigm for decentralized model training, enabling multiple clients to collaboratively lea

Predicting Where Steering Vectors Succeed

Model ReleasesDGX agent

arXiv:2604.15557v1 Announce Type: cross Abstract: Steering vectors work for some concepts and layers but fail for others, and practitioners have no way to predict which setting applies before running

RAGognizer: Hallucination-Aware Fine-Tuning via Detection Head Integration

ResearchDGX agent

arXiv:2604.15945v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) is widely used to augment the input to Large Language Models (LLMs) with external information, such as recent or do

ReactBench: A Benchmark for Topological Reasoning in MLLMs on Chemical Reaction Diagrams

Model ReleasesDGX agent

arXiv:2604.15994v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at recognizing individual visual elements and reasoning over simple linear diagrams. However, when faced

RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity

Model ReleasesDGX agent

arXiv:2509.25897v2 Announce Type: replace-cross Abstract: People often encounter role conflicts -- social dilemmas where the expectations of multiple roles clash and cannot be simultaneously fulfilled

SecureRouter: Encrypted Routing for Efficient Secure Inference

ApplicationsDGX agent

arXiv:2604.15499v1 Announce Type: cross Abstract: Cryptographically secure neural network inference typically relies on secure computing techniques such as Secure Multi-Party Computation (MPC), enabli

Self-Aligned Reward: Towards Effective and Efficient Reasoners

ResearchDGX agent

arXiv:2509.05489v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards has significantly advanced reasoning in large language models (LLMs), but such signals remain coarse,

Structured Abductive-Deductive-Inductive Reasoning for LLMs via Algebraic Invariants

ResearchDGX agent

arXiv:2604.15727v1 Announce Type: new Abstract: Large language models exhibit systematic limitations in structured logical reasoning: they conflate hypothesis generation with verification, cannot dist

Subjective and Objective Quality-of-Experience Evaluation Study for Live Video Streaming

Model ReleasesDGX agent

arXiv:2409.17596v2 Announce Type: replace-cross Abstract: In recent years, live video streaming has gained widespread popularity across various social media platforms. Quality of experience (QoE), whi

The Reasoning Trap: How Enhancing LLM Reasoning Amplifies Tool Hallucination

Model ReleasesDGX agent

arXiv:2510.22977v2 Announce Type: replace-cross Abstract: Enhancing the reasoning capabilities of Large Language Models (LLMs) is a key strategy for building Agents that 'think then act.' However, rec

Towards Rigorous Explainability by Feature Attribution

ResearchDGX agent

arXiv:2604.15898v1 Announce Type: new Abstract: For around a decade, non-symbolic methods have been the option of choice when explaining complex machine learning (ML) models. Unfortunately, such metho

Training Flow Matching: The Role of Weighting and Parameterization

ResearchDGX agent

arXiv:2603.06454v2 Announce Type: replace Abstract: We study the training objectives of denoising-based generative models, with a particular focus on loss weighting and output parameterization, includ

Watching Movies Like a Human: Egocentric Emotion Understanding for Embodied Companions

Model ReleasesDGX agent

arXiv:2604.15823v1 Announce Type: new Abstract: Embodied robotic agents often perceive movies through an egocentric screen-view interface rather than native cinematic footage, introducing domain shift

WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback

SafetyDGX agent

arXiv:2408.15549v4 Announce Type: replace Abstract: As large language models (LLMs) continue to advance, aligning these models with human preferences has emerged as a critical challenge. Traditional a

Zoom Consistency: A Free Confidence Signal in Multi-Step Visual Grounding Pipelines

ResearchDGX agent

arXiv:2604.15376v1 Announce Type: cross Abstract: Multi-step zoom-in pipelines are widely used for GUI grounding, yet the intermediate predictions they produce are typically discarded after coordinate

19 Apr 2026

“Salaryman eating ramen” is like the Eastern equivalent of “Will Smith eating spaghetti” test

ResearchDGX agent

This post draws a humorous parallel between using images of salarymen eating ramen as a test case for AI image generation models in Eastern contexts and the Western 'Will Smith eating spaghetti' meme,

Since Anthropic publish their system prompts we can generate a diff between Claude Opus 4.6 and 4.7 - here are my notes on what's changed ht…

Model ReleasesDGX agent

Simon Willison documents the differences between Anthropic's Claude Opus 4.6 and 4.7 system prompts, analyzing changes that Anthropic made public. The notes likely highlight modifications to model beh

18 Apr 2026

Sweet spot…Cloud & local LLM setup + Mission Control

Local AiDGX agent

A discussion exploring the 'sweet spot' of using Ollama's hybrid Cloud + Local setup, where a reachable Ollama host serves as the control point for both local and cloud models . The post likely covers

17 Apr 2026

3D Instruction Ambiguity Detection

Model ReleasesDGX agent

arXiv:2601.05991v2 Announce Type: replace Abstract: In safety-critical domains, linguistic ambiguity can have severe consequences; a vague command like 'Pass me the vial' in a surgical setting could l

Active Learning with Selective Time-Step Acquisition for PDEs

Model ReleasesDGX agent

arXiv:2511.18107v2 Announce Type: replace Abstract: Accurately solving partial differential equations (PDEs) is critical to understanding complex scientific and engineering phenomena, yet traditional

AI-Assisted Peer Review at Scale: The AAAI-26 AI Review Pilot

Model ReleasesDGX agent

arXiv:2604.13940v1 Announce Type: new Abstract: Scientific peer review faces mounting strain as submission volumes surge, making it increasingly difficult to sustain review quality, consistency, and t

[AINews] Anthropic Claude Opus 4.7 - literally one step better than 4.6 in every dimension

Model ReleasesDGX agent

This article from Latent Space discusses Anthropic's Claude Opus 4.7 release, highlighting incremental improvements across multiple performance dimensions compared to the previous 4.6 version. The pie

Applying an Agentic Coding Tool for Improving Published Algorithm Implementations

Model ReleasesDGX agent

arXiv:2604.13109v1 Announce Type: cross Abstract: We present a two-stage pipeline for AI-assisted improvement of published algorithm implementations. In the first stage, a large language model with re

ArrowGEV: Grounding Events in Video via Learning the Arrow of Time

ResearchDGX agent

arXiv:2601.06559v2 Announce Type: replace Abstract: Grounding events in videos serves as a fundamental capability in video analysis. While Vision Language Models (VLMs) are increasingly employed for t

Assessment Design in the AI Era: A Method for Identifying Items Functioning Differentially for Humans and Chatbots

Model ReleasesDGX agent

arXiv:2603.23682v2 Announce Type: replace-cross Abstract: The rapid adoption of large language models (LLMs) in education raises profound challenges for assessment design. To adapt assessments to the

ChatSVA: Bridging SVA Generation for Hardware Verification via Task-Specific LLMs

Model ReleasesDGX agent

arXiv:2604.02811v2 Announce Type: replace-cross Abstract: Functional verification consumes over 50% of the IC development lifecycle, where SystemVerilog Assertions (SVAs) are indispensable for formal

Class Unlearning via Depth-Aware Removal of Forget-Specific Directions

Model ReleasesDGX agent

arXiv:2604.15166v1 Announce Type: new Abstract: Machine unlearning aims to remove targeted knowledge from a trained model without the cost of retraining from scratch. In class unlearning, however, red

Comparison of Modern Multilingual Text Embedding Techniques for Hate Speech Detection Task

Model ReleasesDGX agent

arXiv:2604.14907v1 Announce Type: new Abstract: Online hate speech and abusive language pose a growing challenge for content moderation, especially in multilingual settings and for low-resource langua

Create Expert Content: Deploying a Multi-Agent System with Terraform and Cloud Run

Model ReleasesDGX agent

In support of our mission to accelerate the developer journey on Google Cloud, we built Dev Signal: a multi-agent system designed to transform raw community signals into reliable technical guidance by

DocVAL: Validated Chain-of-Thought Distillation for Grounded Document VQA

Local AiDGX agent

arXiv:2511.22521v2 Announce Type: replace Abstract: Document visual question answering requires models not only to answer questions correctly, but also to precisely localize answers within complex doc

Don't Retrieve, Navigate: Distilling Enterprise Knowledge into Navigable Agent Skills for QA and RAG

Model ReleasesDGX agent

arXiv:2604.14572v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) grounds LLM responses in external evidence but treats the model as a passive consumer of search results: it never

Evolving Media CDN for the world’s most demanding broadcast and streaming workloads

Model ReleasesDGX agent

Editor’s note: In this post, we share joint insights from Raj Gulani, Director of Product Management for Network Experiences, and Dan Rayburn, Industry analyst with 30-plus years of experience coverin

Federated Breast Cancer Detection Enhanced by Synthetic Ultrasound Image Augmentation

ResearchDGX agent

arXiv:2506.23334v3 Announce Type: replace-cross Abstract: Federated learning enables collaborative training of deep learning models across institutions without sharing sensitive patient data. However,

Federated Multi-Task Clustering

TutorialsDGX agent

arXiv:2512.22897v3 Announce Type: replace Abstract: Spectral clustering has emerged as one of the most effective clustering algorithms due to its superior performance. However, most existing models ar

From Memorization to Creativity: LLM as a Designer of Novel Neural Architectures

Model ReleasesDGX agent

arXiv:2601.02997v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel in program synthesis, yet their capacity for neural architecture design -- balancing syntactic reliability,

Graph-Based Fraud Detection with Dual-Path Graph Filtering

ApplicationsDGX agent

arXiv:2604.14235v1 Announce Type: new Abstract: Fraud detection on graph data can be viewed as a demanding task that requires distinguishing between different types of nodes. Because graph neural netw

← Previous
1…428429430431432…1059
Next →