Bounded knowledge map
Recent connections
Explicit wiki links among recent Model Releases entries.
90 nodes · 3 links. The server bounds this neighbourhood, so the browser never downloads or simulates the complete wiki graph.
90 nodes · 3 links shown. Hover or focus to isolate neighbours; select a node to open its source entry.
Linked entries
3 shownPower law graph attention: exact generalization of scaled dot-product attention, empirical collapse at inference→UT-ACA: Uncertainty-Triggered Adaptive Context Allocation for Long-Context Inference
Model Releases
90 shownPower law graph attention: exact generalization of scaled dot-product attention, empirical collapse at inference1UT-ACA: Uncertainty-Triggered Adaptive Context Allocation for Long-Context Inference1Seq2Synth: Benchmarking Temporal Fidelity in Synthetic Sequential Tabular Data1Benchmarking Time Series Generation Methods for Privacy-Preserving Forecasting1Test-Time Self-Evolving GUI Visual Grounding via Reflection-Guided On-Policy Self-Distillation1SkillLens: Visual Skill Cards for Retrieval-Augmented GUI Action Prediction and On-Policy Distillation1FormStruct-Bench:A Hierarchical and Diagnostic Benchmark for Table-Form Document Structure RecognitionReasoning Shortcuts and Value Symmetries: What Symmetry Permits, Architecture Realizes, and Optimization SelectsMeasure the Sim-to-Real Gap: Designing an Affordable Real-World Benchmark Platform for Reinforcement Learning in AIoT SystemsNutrition Data Infrastructure for the AI Era: Operationalizing FAIR for Agent-Mediated ResearchREDAgentBench: Executable Red Teaming and Faithful Measurement of LLM Agent SystemsIdea for a deepseek-v4-flash-0731 backed automated research workflow to be leveraged via qwen3.6/3.8 27b for difficult tasks that require highly technical, not easy to find information.The Truth Stays in the Family: Enhancing Contextual Grounding via Inherited Truthful Heads in Model LineagesTEASR: Training-Efficient Any-Step Diffusion Transformer for Real-World Image Super-ResolutionVibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?A variational Bayes approach to inference for low-dimensional parameters in high-dimensional linear regressionImplicit representations are dead. Long live explicit primitives!Tested Nemotron 3.5 Lightning locally on coding, Hermes Agent and agentic workENTLORE: A Graph-Grounded Benchmark for Latent Organizational Reasoning in Enterprise Question AnsweringFlex-pi: A Multi-Stream World-Action Model with Compute FlexibilityCost-Efficient Estimation of General Abilities Across BenchmarksReRound: Reconstructive Rounding to Resolve Midpoint Ambiguity in Calibration-Free LLM Quantizationb10375InSight-doc: Agentic Visual Perception for Long-Document UnderstandingTemporally Grounded Compositional Camera Motion Understanding via Geometric Knowledge DistillationDegradeQuery: Counterfactual Tuple Pretraining for Context-Aware PROTAC Degradation PredictionActionable Hallucination Detection: Translating Latent Uncertainty into Agentic Critiqueb10373We have her dash camera which shows they are lying. The agents are wearing body cameras and should have dash cameras of their own. If what t…myMediWhisper: Construction of Burmese Medical Speech Corpus and Whisper Fine-Tuning for Clinical Dialogue ASRSignificance and Stability Analysis of Gene-Environment Interaction using GxEStatNavigation Alone Is Not Enough: Evaluating Explanatory Assistive UI AgentsHSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language ModelsDo LLM Recommenders Know When They're Hallucinating? Auditing Confidence Calibration in Catalog FaithfulnessHoosierHelp: Benchmarking LLM Agents for Social Service NavigationRethinking LLM Verification: Evidence Structure, Uncertainty, and Selective RefinementTowards Efficient Reasoning in LLM-Based Recommender Systems via Model MergingDistilVDR: A Compact End-to-End Visual Document Retriever via Dual-Student DistillationUserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMsEveryone else is talking about building ASI to like monopolize b2b saas and Elon is talking about building a kardashev II sentient sunWhen Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language ModelsIs There Really a Camouflaged Object? Towards Realistic Camouflaged Object DetectionFaithformBench: Benchmarking Faithfulness of Mathematical Chain-of-Thought AutoformalisationMapping and Measuring the Behavioral Evolution of Large Language ModelsPutting sign language AI into users’ handsNeural Introspection Gating for Adaptive KV-Cache Reuse in Vision-Language-Action ModelsOn Solomonoff Induction in Large Language Models and the Limits of Self-Improving: The Singularity Is Not Near Without Symbolic Model SynthesisE^3mo-Bench: A Scalable Benchmark for Multimodal Evoked and Expressed Emotion Understanding via Bayesian Pairwise AlignmentStream Forcing: Constructing Unified Training Trajectory for Robust Streaming Video GenerationExploration-Driven Personalized Federated Reinforcement Learning via Intrinsic MotivationMeasuring Semantic Abstractness of SAE Features via Nonlocalitypi-SUB: A Physics-Informed Synthetic Underwater Benchmark Dataset for Underwater Image EnhancementGeoSeg-OV: Bridging Geospatial Gaps with Structural Guidance for Open-Vocabulary Remote Sensing SegmentationBenchmarking LLM-Guided Control-Plane Policies for Backend Fault Isolation in HAProxyTACTICL: Task-Aware Compression of Tabular ICL ModelsCan Released LLM Vocabularies Support Token-Level Estimation of Hidden Corpora?Putting Registers to Work: Task Registers for Token Pruning in Vision TransformersEdge Phoneme Recognition for Children's Speech through Age-Aware Trainingb10369Derivative Computation in PINNs: Automatic Differentiation, Finite Differences and BeyondNo Free Labels: Limitations of LLM-as-a-Judge Without Human GroundingAhrefs launches AI agent workspace Letaido for marketers and agenciesMAD-HOI: Masked Autoregressive Diffusion for Generating Articulated Hand Object Interactions from TextEvidence-Grounded Trustworthy Multimodal Reasoning and Evaluation Benchmark in Complex Urban ScenesOn the Limitations of Cross-Lingual Consistency in Multilingual Text-to-image GenerationDiffract: Spectral View of LLM Domain AdaptationOptimized Sequential Testing for Binary Ensemble ClassifiersGESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic ScenesSeFoRA: Sketch-Aggregated Federated Low-Rank Adaptation with Heterogeneous Client RanksV-FiLLM: Verified Financial LLM Reasoning BenchmarkIs This Your Final Answer? Cross-Contextual Consistency as a Measure of LLM CredibilityTowards Unified Dynamic Face Landmark DetectionVision-Language-Motion Maps: An Open-Vocabulary, Uncertainty-Aware, Queryable Motion Attribute for 3D Scene MapsBridging Severe Cross-Modal Misalignment: End-to-End Visible-Infrared Object Detection via Explicit Feature-Domain Affine RegistrationComBodied Agents: a New Paradigm of Human-Centric Agentic AIReLTEx: Reliable LLM-based Taxonomy ExpansionCurate Before You Connect: Identity and Ontology Tagging in a Production Knowledge GraphAIFS-TC: A simple correction competitive with the operational frontier for tropical cyclone intensity forecastingFast and Memory-Efficient Wavelet Convolutions via I/O-Aware ReformulationTemMed-Bench: Evaluating Temporal Medical Image Reasoning in Vision-Language ModelsStatic in Frames, Dynamic in Events: Rethinking Features in Event Cameras as Motion CuesSAR2Agri: Learning SAR Intensity Representations for Agricultural MonitoringLocally Deployable Small Language Models for Emergency Department Decision Support: A Systematic Benchmark of Fine-Tuning StrategiesA HamNoSys-Guided Dataset and Baselines for Fine-Grained Isolated Handshape Recognition in Sign LanguageWhat unique, custom QOL upgrades have you given your local agents?CHORUS: Complementary Experts for High-Coverage Testbench Stimulus GenerationDACRI: Decision-Aware Causal Intervention Ranking for Critical Supply ChainsKKL Observer Synthesis for Nonlinear Systems via Physics-Informed LearningHyperShape: Hyperelasticity Across Diverse ShapesCracks in the Foundation: Seemingly Minor Architectural Choices Impact Long Context Extension