Slow Kill for Big Data Learning
Big-data applications often involve a vast number of observations and features, creating new challenges for variable selection and parameter estimation. This paper presents a novel technique called ``slow kill,'' which utilizes nonconvex constrained optimization, adaptive $\ell_2$-shrinkage, and increasing learning rates. The fact that the problem size can decrease during the slow kill iterations makes it particularly effective for large-scale variable screening. The interaction between statistics and optimization provides valuable insights into controlling quantiles, stepsize, and shrinkage parameters in order to relax the regularity conditions required to achieve the desired level of statistical accuracy. Experimental results on real and synthetic data show that slow kill outperforms state-of-the-art algorithms in various situations while being computationally efficient for large-scale data.
Code (0)
등록된 구현이 없습니다.
Tasks
parameter estimationVariable SelectionSimilar Papers 제목 키워드 기반
Multidimensional Skills on LinkedIn Profiles: Measuring Human Capital and the Gender Skill Gap
We measure human capital using the self-reported skill sets of nearly 9 million U.S. college graduates from professional profiles on LinkedIn. We aggregate skill strings into 48 clusters of general, occupation-specific, …
EffiSkill: Agent Skill Based Automated Code Efficiency Optimization
Code efficiency is a fundamental aspect of software quality, yet how to harness large language models (LLMs) to optimize programs remains challenging. Prior approaches have sought for one-shot rewriting, retrieved exempl…
Dual-Process Atomic Skill Learning: Decoupling Semantic Reasoning and Real-Time Control
Language-conditioned Imitation Learning (IL) is essential for enabling robots to perform complex tasks following natural language instructions. However, generalizing to multi-step compositional tasks remains a significan…
SkillGPT: a RESTful API service for skill extraction and standardization using a Large Language Model
We present SkillGPT, a tool for skill extraction and standardization (SES) from free-style job descriptions and user profiles with an open-source Large Language Model (LLM) as backbone. Most previous methods for similar …
Feature EngineeringLanguage ModelingLanguage ModellingLarge Language ModelHigh-level Features for Resource Economy and Fast Learning in Skill Transfer
Abstraction is an important aspect of intelligence which enables agents to construct robust representations for effective decision making. In the last decade, deep networks are proven to be effective due to their ability…
Decision Making