PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 31, 2012ACM Transactions on Knowledge Discovery from Data71 citationsOpen Access

Large Linear Classification When Data Cannot Fit in Memory

HYHsiang‐Fu YuCHCho‐Jui HsiehKCKai‐Wei Chang

Key Points

Key points are not available for this paper at this time.

Abstract

Recent advances in linear classification have shown that for applications such as document classification, the training process can be extremely efficient. However, most of the existing training methods are designed by assuming that data can be stored in the computer memory. These methods cannot be easily applied to data larger than the memory capacity due to the random access to the disk. We propose and analyze a block minimization framework for data larger than the memory size. At each step a block of data is loaded from the disk and handled by certain learning methods. We investigate two implementations of the proposed framework for primal and dual SVMs, respectively. Because data cannot fit in memory, many design considerations are very different from those for traditional algorithms. We discuss and compare with existing approaches that are able to handle data larger than memory. Experiments using data sets 20 times larger than the memory demonstrate the effectiveness of the proposed method.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Yu et al. (2012) studied this question.

synapsesocial.com/papers/6a1c049bb33628da419d1e1dhttps://doi.org/10.1145/2086737.2086743
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1On the convergence of the coordinate descent method for convex differentiable minimization1992 · 519 citations
  2. 2LIBSVM2011 · 41,487 citations
  3. 3Classifying large data sets using SVMs with hierarchical clusters2003 · 275 citations
  4. 4Interior-Point Methods for Massive Support Vector Machines2002 · 190 citations
  5. 5Slow Learners are Fast2009 · 201 citations