Lv31
360 积分 2024-01-16 加入
Cross-Platform Evaluation of Reasoning Capabilities in Foundation Models
2小时前
已完结
The Impact of Parameter Scaling: Analysis of Specific Large Language Model Capabilities
2小时前
已完结
Sloth: scaling laws for LLM skills to predict multi-benchmark performance across families
2小时前
已完结
Sloth: scaling laws for LLM skills to predict multi-benchmark performance across families
2小时前
已完结
LLM-as-a-Prophet: Understanding Predictive Intelligence with Prophet Arena
2小时前
已完结
Forecasting Frontier Language Model Agent Capabilities
2小时前
已完结
SampleMix: A Sample-wise Pre-training Data Mixing Strategey by Coordinating Data Quality and Diversity
2小时前
已完结
An Information Theory of Compute-Optimal Size Scaling, Emergence, and Plateaus in Language Models
2小时前
已完结
Capability Salience Vector: Fine-grained Alignment of Loss and Capabilities for Downstream Task Scaling Law
2小时前
已完结
Data Mixing for Large Language Models Pretraining: A Survey and Outlook
2小时前
已完结