AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,553 results
21 Apr 2026

Efficient Low-Resource Language Adaptation via Multi-Source Dynamic Logit Fusion

ResearchDGX agent

arXiv:2604.18106v1 Announce Type: new Abstract: Adapting large language models (LLMs) to low-resource languages (LRLs) is constrained by the scarcity of task data and computational resources. Although

Fairness Constraints in High-Dimensional Generalized Linear Models

SafetyDGX agent

arXiv:2604.16610v1 Announce Type: cross Abstract: Machine learning models often inherit biases from historical data, raising critical concerns about fairness and accountability. Conventional fairness

How Tokenization Limits Phonological Knowledge Representation in Language Models and How to Improve Them

Local AiDGX agent

arXiv:2604.17105v1 Announce Type: new Abstract: Tokenization is the first step in every language model (LM), yet it never takes the sounds of words into account. We investigate how tokenization influe

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Kimi 2.6 Thinking seems very good for an open weights model, but many rough edges compared to closed SoTA. The Lem Test resulted in a 74 pag…

ApplicationsDGX agent

Kimi 2.6 Thinking seems very good for an open weights model, but many rough edges compared to closed SoTA. The Lem Test resulted in a 74 page thinking trace... and an okay-ish answer. It did an okay T

Kimi K2.6 has captured #1 on the open-weight Vals Index, and is #7 overall.

Model ReleasesDGX agent

Kimi K2.6, a language model developed by Moonshot AI, has achieved the top ranking on the open-weight category of the Vals Index benchmark, while placing 7th overall across all model categories. The V

Learning to Trade Like an Expert: Cognitive Fine-Tuning for Stable Financial Reasoning in Language Models

AgentsDGX agent

arXiv:2604.16862v1 Announce Type: new Abstract: Recent deployments of large language models (LLMs) as autonomous trading agents raise questions about whether financial decision-making competence gener

Leveraging Large Language Models for Sarcastic Speech Annotation in Sarcasm Detection

Model ReleasesDGX agent

arXiv:2506.00955v2 Announce Type: replace Abstract: Sarcasm fundamentally alters meaning through tone and context, yet detecting it in speech remains a challenge due to data scarcity. In addition, exi

LLM-AUG: Robust Wireless Data Augmentation with In-Context Learning in Large Language Models

ResearchDGX agent

arXiv:2604.17770v1 Announce Type: new Abstract: Data scarcity remains a fundamental bottleneck in applying deep learning to wireless communication problems, particularly in scenarios where collecting

Low-rank Orthogonalization for Large-scale Matrix Optimization with Applications to Foundation Model Training

Model ReleasesDGX agent

arXiv:2509.11983v2 Announce Type: replace Abstract: Neural network (NN) training is inherently a large-scale matrix optimization problem, yet the matrix structure of NN parameters has long been overlo

Mastering the 600B+ Frontier: Optimizing Large Model Deployments on the Inference Cloud

IndustryDGX agent

This DigitalOcean guide covers strategies and best practices for deploying and optimizing very large language models (600 billion+ parameters) on cloud infrastructure, focusing on inference performanc

MeasHalu: Mitigation of Scientific Measurement Hallucinations for Large Language Models with Enhanced Reasoning

Model ReleasesDGX agent

arXiv:2604.16929v1 Announce Type: new Abstract: The accurate extraction of scientific measurements from literature is a critical yet challenging task in AI4Science, enabling large-scale analysis and i

MHSafeEval: Role-Aware Interaction-Level Evaluation of Mental Health Safety in Large Language Models

SafetyDGX agent

arXiv:2604.17730v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly explored as scalable tools for mental health counseling, yet evaluating their safety remains challenging d

Modeling Human Perspectives with Socio-Demographic Representations

ApplicationsDGX agent

arXiv:2604.18069v1 Announce Type: new Abstract: Humans often hold different perspectives on the same issues. In many NLP tasks, annotation disagreement can reflect valid subjective perspectives. Model

Modeling Multiple Support Strategies within a Single Turn for Emotional Support Conversations

Model ReleasesDGX agent

arXiv:2604.17972v1 Announce Type: new Abstract: Emotional Support Conversation (ESC) aims to assist individuals experiencing distress by generating empathetic and supportive dialogue. While prior work

NaviFormer: A Deep Reinforcement Learning Transformer-like Model to Holistically Solve the Navigation Problem

ApplicationsDGX agent

arXiv:2604.16967v1 Announce Type: new Abstract: Path planning is usually solved by addressing either the (high-level) route planning problem (waypoint sequencing to achieve the final goal) or the (low

Neural Garbage Collection: Learning to Forget while Learning to Reason

TutorialsDGX agent

arXiv:2604.18002v1 Announce Type: new Abstract: Chain-of-thought reasoning has driven striking advances in language model capability, yet every reasoning step grows the KV cache, creating a bottleneck

On Different Notions of Redundancy in Conditional-Independence-Based Discovery of Graphical Models

ResearchDGX agent

arXiv:2502.08531v3 Announce Type: replace Abstract: Conditional-independence-based discovery uses statistical tests to identify a graphical model that represents the independence structure of variable

On the Interpolation Effect of Score Smoothing in Diffusion Models

ResearchDGX agent

arXiv:2502.19499v3 Announce Type: replace Abstract: Diffusion models have achieved remarkable progress in various domains with an intriguing ability to produce new data that do not exist in the traini

Precise Debugging Benchmark: Is Your Model Debugging or Regenerating?

Model ReleasesDGX agent

arXiv:2604.17338v1 Announce Type: cross Abstract: Unlike code completion, debugging requires localizing faults and applying targeted edits. We observe that frontier LLMs often regenerate correct but o

Reward Score Matching: Unifying Reward-based Fine-tuning for Flow and Diffusion Models

SafetyDGX agent

arXiv:2604.17415v1 Announce Type: cross Abstract: Reward-based fine-tuning aims to steer a pretrained diffusion or flow-based generative model toward higher-reward samples while remaining close to the

SafeLM: Unified Privacy-Aware Optimization for Trustworthy Federated Large Language Models

SafetyDGX agent

arXiv:2604.16606v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in high-stakes domains, yet a unified treatment of their overlapping safety challenges remains

Safety, Security, and Cognitive Risks in State-Space Models: A Systematic Threat Analysis with Spectral, Stateful, and Capacity Attacks

SafetyDGX agent

arXiv:2604.16424v1 Announce Type: cross Abstract: State-Space Models (SSMs) -- structured SSMs (S4, S4D, DSS, S5), selective SSMs (Mamba, Mamba-2), and hybrid architectures (Jamba) -- are deployed in

Sparse Feature Coactivation Reveals Causal Semantic Modules in Large Language Models

ResearchDGX agent

arXiv:2506.18141v3 Announce Type: replace Abstract: We identify semantically coherent, context-consistent network components in large language models (LLMs) using coactivation of sparse autoencoder (S

StageMem: Lifecycle-Managed Memory for Language Models

ResearchDGX agent

arXiv:2604.16774v1 Announce Type: new Abstract: Long-horizon language model systems increasingly rely on persistent memory, yet many current designs still treat memory primarily as a static store: wri

Synthetic Data Generation for Training Diversified Commonsense Reasoning Models

ResearchDGX agent

arXiv:2603.18361v2 Announce Type: replace Abstract: Conversational agents are required to respond to their users not only with high quality (i.e. commonsense bearing) responses, but also considering m

True to form, I've already seen OpenAI themselves refer to the new image model as 'ChatGPT Images 2.0', 'Image gen 2' and 'gpt-image-2'

ToolsDGX agent

OpenAI has been using multiple informal names internally and externally for its new image generation model, including 'ChatGPT Images 2.0,' 'Image gen 2,' and 'gpt-image-2.' The post highlights incons

VIBE: Voice-Induced open-ended Bias Evaluation for Large Audio-Language Models via Real-World Speech

SafetyDGX agent

arXiv:2604.17248v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) are increasingly integrated into daily applications, yet their generative biases remain underexplored. Existing sp

When Earth Foundation Models Meet Diffusion: An Application to Land Surface Temperature Super-Resolution

Model ReleasesDGX agent

arXiv:2604.16841v1 Announce Type: new Abstract: Land surface temperature (LST) super-resolution is important for environmental monitoring. However, it remains challenging as coarse thermal observation

When Visuals Aren't the Problem: Evaluating Vision-Language Models on Misleading Data Visualizations

Model ReleasesDGX agent

arXiv:2603.22368v2 Announce Type: replace Abstract: Visualizations help communicate data insights, but deceptive data representations can distort their interpretation and propagate misinformation. Whi

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments

AgentsDGX agent

arXiv:2604.18484v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models drive next-generation autonomous systems, but training them requires scalable, high-quality annotations from complex

20 Apr 2026

A Systematic Study of Training-Free Methods for Trustworthy Large Language Models

SafetyDGX agent

arXiv:2604.15789v1 Announce Type: new Abstract: As Large Language Models (LLMs) receive increasing attention and are being deployed across various domains, their potential risks, including generating

Adapting in the Dark: Efficient and Stable Test-Time Adaptation for Black-Box Models

Local AiDGX agent

arXiv:2604.15609v1 Announce Type: cross Abstract: Test-Time Adaptation (TTA) for black-box models accessible only via APIs remains a largely unexplored challenge. Existing approaches such as post-hoc

Anthropic's Mythos AI model sparks fears of turbocharged hacking

IndustryDGX agent

Anthropic announced its Mythos AI model on April 7 and refused to release it publicly, citing its unprecedented ability to find and exploit software vulnerabilities . The company is instead limiting a

Characterising LLM-Generated Competency Questions: a Cross-Domain Empirical Study using Open and Closed Models

Model ReleasesDGX agent

arXiv:2604.16258v1 Announce Type: new Abstract: Competency Questions (CQs) are a cornerstone of requirement elicitation in ontology engineering. CQs represent requirements as a set of natural language

CoMeT: Collaborative Memory Transformer for Efficient Long Context Modeling

Model ReleasesDGX agent

arXiv:2602.01766v2 Announce Type: replace-cross Abstract: The quadratic complexity and indefinitely growing key-value (KV) cache of standard Transformers pose a major barrier to long-context processin

DiZiNER: Disagreement-guided Instruction Refinement via Pilot Annotation Simulation for Zero-shot Named Entity Recognition

Model ReleasesDGX agent

arXiv:2604.15866v1 Announce Type: cross Abstract: Large language models (LLMs) have advanced information extraction (IE) by enabling zero-shot and few-shot named entity recognition (NER), yet their ge

EchoVLM: Dynamic Mixture-of-Experts Vision-Language Model for Universal Ultrasound Intelligence

ResearchDGX agent

arXiv:2509.14977v2 Announce Type: replace Abstract: Ultrasound imaging has become the preferred imaging modality for early cancer screening due to its advantages of non-ionizing radiation, low cost, a

Kimi has been the most popular model on Fireworks, both out of the box and as a fine-tuning base (including Composer 2) Now Kimi K2.6 is liv…

ToolsDGX agent

Kimi has been the most popular model on Fireworks, both out of the box and as a fine-tuning base (including Composer 2) Now Kimi K2.6 is live with huge jumps (10+%) in coding, long-running agents and

KWBench: Measuring Unprompted Problem Recognition in Knowledge Work

Model ReleasesDGX agent

arXiv:2604.15760v1 Announce Type: new Abstract: We introduce the first version of KWBench (Knowledge Work Bench), a benchmark for unprompted problem recognition in large language models: can an LLM id

Large Language Models for Market Research: A Data-augmentation Approach

SafetyDGX agent

arXiv:2412.19363v3 Announce Type: replace Abstract: Large Language Models (LLMs) have transformed artificial intelligence by excelling in complex natural language processing tasks. Their ability to ge

Large Reasoning Models Are (Not Yet) Multilingual Latent Reasoners

ResearchDGX agent

arXiv:2601.02996v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieve strong performance on mathematical reasoning tasks, often attributed to their capability to generate explicit

literally my basic model since 1998. crazy that some people still haven’t figured this out.

SafetyDGX agent

literally my basic model since 1998. crazy that some people still haven’t figured this out. My basic model of capabilities: LLMs are good at problems similar to those that appear in their training dat

Modeling of ASD/TD Children's Behaviors in Interaction with a Virtual Social Robot During a Music Education Program Using Deep Neural Networks

ApplicationsDGX agent

arXiv:2604.15314v1 Announce Type: cross Abstract: This research aimed to develop an intelligent system to evaluate performance and extract behavioral models for children with ASD and neurotypical (TD)

Nice paper combining the strength of Skills and RAG. Most RAG systems retrieve on every query, whether the model needs help or not. This is …

AgentsDGX agent

Nice paper combining the strength of Skills and RAG. Most RAG systems retrieve on every query, whether the model needs help or not. This is wasteful when the model already knows the answer, and often

Olmo Hybrid: From Theory to Practice and Back

Model ReleasesDGX agent

arXiv:2604.03444v3 Announce Type: replace-cross Abstract: Recent work has demonstrated the potential of non-transformer language models, especially linear recurrent neural networks (RNNs) and hybrid m

Opportunities and Challenges of Large Language Models for Low-Resource Languages in Humanities Research

ResearchDGX agent

arXiv:2412.04497v5 Announce Type: replace-cross Abstract: Low-resource languages serve as invaluable repositories of human history, embodying cultural evolution and intellectual diversity. Despite the

opus 4.7 dropped last week with two new features: xhigh effort: more compute when you need it task budgets: the model knows its limit and wr…

AgentsDGX agent

opus 4.7 dropped last week with two new features: xhigh effort: more compute when you need it task budgets: the model knows its limit and wraps up gracefully, this is a great example of giving the mod

Our researchers are heading to ICLR with new work: model efficiency, long-context reasoning, next-gen attention and decoding, and more. Chec…

ToolsDGX agent

Our researchers are heading to ICLR with new work: model efficiency, long-context reasoning, next-gen attention and decoding, and more. Check out what we've been building 👇🏼 #TogetherResearch #AINativ

OXtal: An All-Atom Diffusion Model for Organic Crystal Structure Prediction

Model ReleasesDGX agent

arXiv:2512.06987v2 Announce Type: replace Abstract: Accurately predicting experimentally realizable 3D molecular crystal structures from their 2D chemical graphs is a long-standing open challenge in c

{pi}_{0.7}: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

AgentsDGX agent

arXiv:2604.15483v1 Announce Type: new Abstract: We present a new robotic foundation model, called {pi}_{0.7}, that can enable strong out-of-the-box performance in a wide range of scenarios. {pi}_{0.7}

PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation

Model ReleasesDGX agent

arXiv:2604.15670v1 Announce Type: new Abstract: Reasoning segmentation has recently expanded from ground-level scenes to remote-sensing imagery, yet UAV data poses distinct challenges, including obliq

Pruning Unsafe Tickets: A Resource-Efficient Framework for Safer and More Robust LLMs

Model ReleasesDGX agent

arXiv:2604.15780v1 Announce Type: cross Abstract: Machine learning models are increasingly deployed in real-world applications, but even aligned models such as Mistral and LLaVA still exhibit unsafe b

Reward Weighted Classifier-Free Guidance as Policy Improvement in Autoregressive Models

SafetyDGX agent

arXiv:2604.15577v1 Announce Type: cross Abstract: Consider an auto-regressive model that produces outputs x (e.g., answers to questions, molecules) each of which can be summarized by an attribute vect

Scalable spatial point process models for forensic footwear analysis

ResearchDGX agent

arXiv:2602.07006v2 Announce Type: replace Abstract: Shoe print evidence recovered from crime scenes plays a key role in forensic investigations. By examining shoe prints, investigators can determine d

SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos

ResearchDGX agent

arXiv:2602.05638v3 Announce Type: replace Abstract: While foundation models have advanced surgical video analysis, current approaches rely predominantly on pixel-level reconstruction objectives that w

Teaching Language Models Mechanistic Explainability Through MechSMILES

ResearchDGX agent

arXiv:2512.05722v2 Announce Type: replace Abstract: Chemical reaction mechanisms are the foundation of how chemists evaluate reactivity and feasibility, yet current Computer-Assisted Synthesis Plannin

The threat of analytic flexibility in using large language models to simulate human data

HardwareDGX agent

arXiv:2509.13397v3 Announce Type: replace-cross Abstract: Social scientists are now using large language models to create 'silicon samples': synthetic datasets intended to stand in for human responden

VIB-Probe: Detecting and Mitigating Hallucinations in Vision-Language Models via Variational Information Bottleneck

ResearchDGX agent

arXiv:2601.05547v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable progress in multimodal tasks, but remain susceptible to hallucinations, where gener

18 Apr 2026

A major lesson to take away from Opus 4.7 is that, while there is a lot of arguments about implementation choices and personality, models ke…

ApplicationsDGX agent

A major lesson to take away from Opus 4.7 is that, while there is a lot of arguments about implementation choices and personality, models keep improving measurably on economically important tasks with

17 Apr 2026

AlphaCNOT: Learning CNOT Minimization with Model-Based Planning

ResearchDGX agent

arXiv:2604.13812v1 Announce Type: new Abstract: Quantum circuit optimization is a central task in Quantum Computing, as current Noisy Intermediate Scale Quantum devices suffer from error propagation t

← Previous
1…142143144145146…1010
Next →