SOTAVerified

Mathematical Proofs

Papers

Showing 1–50 of 90 papers

TitleStatusHype
A New Era in Software Security: Towards Self-Healing Software via Large Language Models and Formal VerificationCode2
ODGS: 3D Scene Reconstruction from Omnidirectional Images with 3D Gaussian SplattingsCode2
Sharpness-Aware Minimization Alone can Improve Adversarial RobustnessCode1
FormalAlign: Automated Alignment Evaluation for AutoformalizationCode1
AdaSwarm: Augmenting Gradient-Based optimizers in Deep Learning with Swarm IntelligenceCode1
IsarStep: a Benchmark for High-level Mathematical ReasoningCode1
Differential Machine LearningCode1
Simple but Effective Compound Geometric Operations for Temporal Knowledge Graph CompletionCode1
Draft, Sketch, and Prove: Guiding Formal Theorem Provers with Informal ProofsCode1
BreastScreening: On the Use of Multi-Modality in Medical Imaging DiagnosisCode1
Theory-guided hard constraint projection (HCP): a knowledge-based data-driven scientific machine learning methodCode1
TransERR: Translation-based Knowledge Graph Embedding via Efficient Relation RotationCode1
SciEx: Benchmarking Large Language Models on Scientific Exams with Human Expert Grading and Automatic GradingCode0
Hierarchical Attention Generates Better ProofsCode0
Towards Autoformalization of Mathematics and Code Correctness: Experiments with Elementary ProofsCode0
Distance-Adaptive Quaternion Knowledge Graph Embedding with Bidirectional RotationCode0
StepProof: Step-by-step verification of natural language mathematical proofsCode0
Learning to Prove Theorems via Interacting with Proof AssistantsCode0
Calibration of P-values for calibration and for deviation of a subpopulation from the full populationCode0
A Theoretical Analysis of Compositional Generalization in Neural Networks: A Necessary and Sufficient ConditionCode0
Mathematical Formalized Problem Solving and Theorem Proving in Different Fields in Lean 4Code0
Learning Rules Explaining Interactive Theorem Proving Tactic PredictionCode0
Formal Development of Safe Automated Driving using Differential Dynamic LogicCode0
Premise Selection for Mathematics by Corpus Analysis and Kernel MethodsCode0
A Unified Parallel Algorithm for Regularized Group PLS Scalable to Big DataCode0
A New Approach Towards AutoformalizationCode0
Epistemic Phase Transitions in Mathematical ProofsCode0
GENTLE: A Genre-Diverse Multilayer Challenge Set for English NLP and Linguistic EvaluationCode0
Examining the impact of forcing function inputs on structural identifiability—0
FastPart: Over-Parameterized Stochastic Gradient Descent for Sparse optimisation on Measures—0
Fast Quasi-Optimal Power Flow of Flexible DC Traction Power Systems—0
Fence Theorem: Preprocessing is Dual-Objective Semantic Structure Isolator in 3D Anomaly Detection—0
Formal Language Knowledge Corpus for Retrieval Augmented Generation—0
Gender Bias of LLM in Economics: An Existentialism Perspective—0
Generating Millions Of Lean Theorems With Proofs By Exploring State Transition Graphs—0
How Analysis Can Teach Us the Optimal Way to Design Neural Operators—0
How Deduction Systems Can Help You To Verify Stability Properties—0
HybridProver: Augmenting Theorem Proving with LLM-Driven Proof Synthesis and Refinement—0
Identification of Probabilities of Causation: A Complete Characterization—0
Interleaver Design for Deep Neural Networks—0
Investor's sentiment in multi-agent model of the continuous double auction—0
Large Language Models' Understanding of Math: Source Criticism and Extrapolation—0
Learning Functions to Study the Benefit of Multitask Learning—0
LemmaHead: RAG Assisted Proof Generation Using Large Language Models—0
Mathematical Approach in Hybrid Beamforming for ISAC Systems—0
On the uncertainty analysis of the data-enabled physics-informed neural network for solving neutron diffusion eigenvalue problem—0
On the uncertainty principle of neural networks—0
Provably safe and human-like car-following behaviors: Part 2. A parsimonious multi-phase model with projected braking—0
Prover Agent: An Agent-based Framework for Formal Mathematical Proofs—0
α-Rank: Multi-Agent Evaluation by Evolution—0
Show:102550
← PrevPage 1 of 2Next →

No leaderboard results yet.