AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Safety

Task Matters: Knowledge Requirements Shape LLM Responses to Context-Memory Conflict

DGX agent

arXiv:2506.06485v4 Announce Type: replace Abstract: Large language models (LLMs) draw on both contextual information and parametric memory, yet these sources can conflict. Prior studies have largely e

safetyarxiv-cs-cl
21 Apr 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

The Impact of Off-Policy Training Data on Probe Generalisation

DGX agent

arXiv:2511.17408v4 Announce Type: replace-cross Abstract: Probing has emerged as a promising method for monitoring large language models (LLMs), enabling cheap inference-time detection of concerning b

safetyarxiv-cs-lg
21 Apr 2026
Safety

The Provenance Gap in Clinical AI: Evidence-Traceable Temporal Knowledge Graphs for Rare Disease Reasoning

DGX agent

arXiv:2604.17114v1 Announce Type: new Abstract: Frontier large language models generate clinically accurate outputs, but their citations are often fabricated. We term this the Provenance Gap. We teste

safetyarxiv-cs-cl
21 Apr 2026
Applications

The Role of Vocabularies in Learning Sparse Representations for Ranking

DGX agent

arXiv:2509.16621v2 Announce Type: replace-cross Abstract: Learned Sparse Retrieval (LSR) such as SPLADE has growing interest for effective semantic 1st stage matching while enjoying the efficiency of

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

TransXion: A High-Fidelity Graph Benchmark for Realistic Anti-Money Laundering

DGX agent

arXiv:2604.17420v1 Announce Type: new Abstract: Money laundering poses severe risks to global financial systems, driving the widespread adoption of machine learning for transaction monitoring. However

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

TriangleMix: Accelerating Prefilling via Decoding-time Contribution Sparsity

DGX agent

arXiv:2507.21526v3 Announce Type: replace Abstract: Large Language Models (LLMs) incur quadratic attention complexity with input length, creating a major time bottleneck in the prefilling stage. Exist

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

TriTS: Time Series Forecasting from a Multimodal Perspective

DGX agent

arXiv:2604.16748v1 Announce Type: new Abstract: Time series forecasting plays a pivotal role in critical sectors such as finance, energy, transportation, and meteorology. However, Long-term Time Serie

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment

DGX agent

arXiv:2405.13068v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have revolutionized various applications, making robust safety alignment essential to prevent harmful outputs. Cu

safetyarxiv-cs-lg
21 Apr 2026
Research

UniMesh: Unifying 3D Mesh Understanding and Generation

DGX agent

arXiv:2604.17472v1 Announce Type: new Abstract: Recent advances in 3D vision have led to specialized models for either 3D understanding (e.g., shape classification, segmentation, reconstruction) or 3D

researcharxiv-cs-cv
21 Apr 2026
Model Releases

VCORE: Variance-Controlled Optimization-based Reweighting for Chain-of-Thought Supervision

DGX agent

arXiv:2510.27462v2 Announce Type: replace Abstract: Supervised fine-tuning (SFT) on long chain-of-thought (CoT) trajectories has emerged as a crucial technique for enhancing the reasoning abilities of

model-releasesarxiv-cs-cl
21 Apr 2026
Research

x1: Learning to Think Adaptively Across Languages and Cultures

DGX agent

arXiv:2604.16917v1 Announce Type: new Abstract: Languages encode distinct abstractions and inductive priors, yet most large language models (LLMs) overlook this diversity by reasoning in a single domi

researcharxiv-cs-cl
21 Apr 2026
Local Ai

A Single Image and Multimodality Is All You Need for Novel View Synthesis

DGX agent

arXiv:2602.17909v2 Announce Type: replace Abstract: Diffusion-based approaches have recently demonstrated strong performance for single-image novel view synthesis by conditioning generative models on

local-aiarxiv-cs-cv
20 Apr 2026
Model Releases

AEGIS: Anchor-Enforced Gradient Isolation for Knowledge-Preserving Vision-Language-Action Fine-Tuning

DGX agent

arXiv:2604.16067v1 Announce Type: cross Abstract: Adapting pre-trained vision-language models (VLMs) for robotic control requires injecting high-magnitude continuous gradients from a flow-matching act

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

An Information-Geometric Approach to Artificial Curiosity

DGX agent

arXiv:2504.06355v2 Announce Type: replace Abstract: Learning in environments with sparse rewards remains a fundamental challenge in reinforcement learning. Artificial curiosity addresses this limitati

model-releasesarxiv-cs-lg
20 Apr 2026
Tutorials

Analyzing Chain of Thought (CoT) Approaches in Control Flow Code Deobfuscation Tasks

DGX agent

arXiv:2604.15390v1 Announce Type: cross Abstract: Code deobfuscation is the task of recovering a readable version of a program while preserving its original behavior. In practice, this often requires

tutorialsarxiv-cs-ai
20 Apr 2026
Research

Beyond Surface Statistics: Robust Conformal Prediction for LLMs via Internal Representations

DGX agent

arXiv:2604.16217v1 Announce Type: cross Abstract: Large language models are increasingly deployed in settings where reliability matters, yet output-level uncertainty signals such as token probabilitie

researcharxiv-cs-ai
20 Apr 2026
Model Releases

Breaking the Training Barrier of Billion-Parameter Universal Machine Learning Interatomic Potentials

DGX agent

arXiv:2604.15821v1 Announce Type: cross Abstract: Universal Machine Learning Interatomic Potentials (uMLIPs), pre-trained on massively diverse datasets encompassing inorganic materials and organic mol

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

Calling all builders 🛠️📣 Google AI Pro and Ultra subscribers will now get increased usage limits and access to Nano Banana Pro and Gemini …

DGX agent

Calling all builders 🛠️📣 Google AI Pro and Ultra subscribers will now get increased usage limits and access to Nano Banana Pro and Gemini Pro models in @GoogleAIStudio — no API key required. Sign in w

model-releasesgoogle-ai--x
20 Apr 2026
Agents

Capture the Flags: Family-Based Evaluation of Agentic LLMs via Semantics-Preserving Transformations

DGX agent

arXiv:2602.05523v2 Announce Type: replace-cross Abstract: Agentic large language models (LLMs) are increasingly evaluated on cybersecurity tasks using capture-the-flag (CTF) benchmarks, yet existing p

agentsarxiv-cs-ai
20 Apr 2026
Model Releases

Consistency Analysis of Sentiment Predictions using Syntactic & Semantic Context Assessment Summarization (SSAS)

DGX agent

arXiv:2604.15547v1 Announce Type: cross Abstract: The fundamental challenge of using Large Language Models (LLMs) for reliable, enterprise-grade analytics, such as sentiment prediction, is the conflic

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Context, from the time of the o1-preview launch: https://www.oneusefulthing.org/p/something-new-on-openais-strawberry

DGX agent

Ethan Mollick shared context about OpenAI's o1-preview model launch, likely discussing the capabilities and implications of this new reasoning-focused AI system that represents a shift toward models d

applicationsethan-mollick--x
20 Apr 2026
Research

CRoCoDiL: Continuous and Robust Conditioned Diffusion for Language

DGX agent

arXiv:2603.20210v3 Announce Type: replace-cross Abstract: Masked Diffusion Models (MDMs) provide an efficient non-causal alternative to autoregressive generation but often struggle with token dependen

researcharxiv-cs-ai
20 Apr 2026
Model Releases

DPrivBench: Benchmarking LLMs' Reasoning for Differential Privacy

DGX agent

arXiv:2604.15851v1 Announce Type: cross Abstract: Differential privacy (DP) has a wide range of applications for protecting data privacy, but designing and verifying DP algorithms requires expert-leve

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

Enhancing AI and Dynamical Subseasonal Forecasts with Probabilistic Bias Correction

DGX agent

arXiv:2604.16238v1 Announce Type: new Abstract: Decision-makers rely on weather forecasts to plant crops, manage wildfires, allocate water and energy, and prepare for weather extremes. Today, such for

safetyarxiv-cs-lg
20 Apr 2026
Model Releases

'Excuse me, may I say something...' CoLabScience, A Proactive AI Assistant for Biomedical Discovery and LLM-Expert Collaborations

DGX agent

arXiv:2604.15588v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into scientific workflows presents exciting opportunities to accelerate biomedical discovery. However,

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Exploring the Capability Boundaries of LLMs in Mastering of Chinese Chouxiang Language

DGX agent

arXiv:2604.15841v1 Announce Type: new Abstract: While large language models (LLMs) have achieved remarkable success in general language tasks, their performance on Chouxiang Language, a representative

model-releasesarxiv-cs-cl
20 Apr 2026
Local Ai

Federated Learning with Quantum Enhanced LSTM for Applications in High Energy Physics

DGX agent

arXiv:2604.15775v1 Announce Type: new Abstract: Learning with large-scale datasets and information-critical applications, such as in High Energy Physics (HEP), demands highly complex, large-scale mode

local-aiarxiv-cs-lg
20 Apr 2026
Model Releases

HiPreNets: High-Precision Neural Networks through Progressive Training

DGX agent

arXiv:2506.15064v3 Announce Type: replace Abstract: Deep neural networks are powerful tools for solving nonlinear problems in science and engineering, but training highly accurate models becomes chall

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

HyCal: A Training-Free Prototype Calibration Method for Cross-Discipline Few-Shot Class-Incremental Learning

DGX agent

arXiv:2604.15678v1 Announce Type: new Abstract: Pretrained Vision-Language Models (VLMs) like CLIP show promise in continual learning, but existing Few-Shot Class-Incremental Learning (FSCIL) methods

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

Is this chart lying to me? Automating the detection of misleading visualizations

DGX agent

arXiv:2508.21675v3 Announce Type: replace Abstract: Misleading visualizations are a potent driver of misinformation on social media and the web. By violating chart design principles, they distort data

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

JFinTEB: Japanese Financial Text Embedding Benchmark

DGX agent

arXiv:2604.15882v1 Announce Type: cross Abstract: We introduce JFinTEB, the first comprehensive benchmark specifically designed for evaluating Japanese financial text embeddings. Existing embedding be

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Kimi K2.6 just dropped. And it crushed Claude Opus 4.6 on SWE-Bench Pro. Kimi K2.6: 58.6 GPT-5.4 xhigh: 57.7 Gemini 3.1 Pro: 54.2 Claude Opu…

DGX agent

Kimi K2.6 just dropped. And it crushed Claude Opus 4.6 on SWE-Bench Pro. Kimi K2.6: 58.6 GPT-5.4 xhigh: 57.7 Gemini 3.1 Pro: 54.2 Claude Opus 4.6: 53.4 An open source Chinese model is now #1 on agenti

model-releasesclem-delangue--x
20 Apr 2026
Model Releases

LaMSUM: Amplifying Voices Against Harassment through LLM Guided Extractive Summarization of User Incident Reports

DGX agent

arXiv:2406.15809v5 Announce Type: replace Abstract: Citizen reporting platforms help the public and authorities stay informed about sexual harassment incidents. However, the high volume of data shared

model-releasesarxiv-cs-cl
20 Apr 2026
Tools

llm-openrouter 0.6

DGX agent

Release: llm-openrouter 0.6 llm openrouter refresh command for refreshing the list of available models without waiting for the cache to expire. I added this feature so I could try Kimi 2.6 on OpenRout

toolssimon-willison
20 Apr 2026
Model Releases

Mamba-SSM with LLM Reasoning for Feature Selection: Faithfulness-Aware Biomarker Discovery

DGX agent

arXiv:2604.14334v2 Announce Type: replace-cross Abstract: Gradient saliency from deep sequence models surfaces candidate biomarkers efficiently, but the resulting gene lists can be contaminated by tis

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Mapping High-Performance Regions in Battery Scheduling across Data Uncertainty, Battery Design, and Planning Horizons

DGX agent

arXiv:2604.15360v1 Announce Type: new Abstract: This study presents a triadic analysis of energy storage operation under multi-stage model predictive control, investigating the interplay between data

researcharxiv-cs-lg
20 Apr 2026
Model Releases

Mind DeepResearch Technical Report

DGX agent

arXiv:2604.14518v2 Announce Type: replace Abstract: We present Mind DeepResearch (MindDR), an efficient multi-agent deep research framework that achieves leading performance with only ~30B-parameter m

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Mind's Eye: A Benchmark of Visual Abstraction, Transformation and Composition for Multimodal LLMs

DGX agent

arXiv:2604.16054v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have achieved impressive progress on vision language benchmarks, yet their capacity for visual cognitive and

model-releasesarxiv-cs-ai
20 Apr 2026
Hardware

NeuroMesh: A Unified Neural Inference Framework for Decentralized Multi-Robot Collaboration

DGX agent

arXiv:2604.15475v1 Announce Type: new Abstract: Deploying learned multi-robot models on heterogeneous robots remains challenging due to hardware heterogeneity, communication constraints, and the lack

hardwarearxiv-cs-ro
20 Apr 2026
Model Releases

Optimizing Korean-Centric LLMs via Token Pruning

DGX agent

arXiv:2604.16235v1 Announce Type: new Abstract: This paper presents a systematic benchmark of state-of-the-art multilingual large language models (LLMs) adapted via token pruning - a compression techn

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Polarization by Default: Auditing Recommendation Bias in LLM-Based Content Curation

DGX agent

arXiv:2604.15937v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed to curate and rank human-created content, yet the nature and structure of their biases in these

model-releasesarxiv-cs-ai
20 Apr 2026
Local Ai

Power to the Clients: Federated Learning in a Dictatorship Setting

DGX agent

arXiv:2510.22149v3 Announce Type: replace-cross Abstract: Federated learning (FL) has emerged as a promising paradigm for decentralized model training, enabling multiple clients to collaboratively lea

local-aiarxiv-cs-ai
20 Apr 2026
Model Releases

Predicting Where Steering Vectors Succeed

DGX agent

arXiv:2604.15557v1 Announce Type: cross Abstract: Steering vectors work for some concepts and layers but fail for others, and practitioners have no way to predict which setting applies before running

model-releasesarxiv-cs-cl
20 Apr 2026
Research

RAGognizer: Hallucination-Aware Fine-Tuning via Detection Head Integration

DGX agent

arXiv:2604.15945v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) is widely used to augment the input to Large Language Models (LLMs) with external information, such as recent or do

researcharxiv-cs-cl
20 Apr 2026
Model Releases

ReactBench: A Benchmark for Topological Reasoning in MLLMs on Chemical Reaction Diagrams

DGX agent

arXiv:2604.15994v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at recognizing individual visual elements and reasoning over simple linear diagrams. However, when faced

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity

DGX agent

arXiv:2509.25897v2 Announce Type: replace-cross Abstract: People often encounter role conflicts -- social dilemmas where the expectations of multiple roles clash and cannot be simultaneously fulfilled

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

SecureRouter: Encrypted Routing for Efficient Secure Inference

DGX agent

arXiv:2604.15499v1 Announce Type: cross Abstract: Cryptographically secure neural network inference typically relies on secure computing techniques such as Secure Multi-Party Computation (MPC), enabli

applicationsarxiv-cs-ai
20 Apr 2026
Research

Self-Aligned Reward: Towards Effective and Efficient Reasoners

DGX agent

arXiv:2509.05489v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards has significantly advanced reasoning in large language models (LLMs), but such signals remain coarse,

researcharxiv-cs-lg
20 Apr 2026
← Previous
1…555556557558559…1371
Next →