Sha256: 0512a5b615ff3c5219d13b941f5caab6fd9ab39e06453d33b0ca40b6cac9f671

Contents?: true

Size: 370 Bytes

Versions: 5

Compression:

Stored size: 370 Bytes

Contents

module Boilerpipe::Extractors
  class NumWordsRulesExtractor
    def self.text(contents)
      doc = ::Boilerpipe::SAX::BoilerpipeHTMLParser.parse(contents)
      ::Boilerpipe::Extractors::NumWordsRulesExtractor.process doc
      doc.content
    end

    def self.process(doc)
      ::Boilerpipe::Filters::NumWordsRulesClassifier.process doc
      doc
    end
  end
end

Version data entries

5 entries across 5 versions & 1 rubygems

Version Path
boilerpipe-ruby-0.5.0 lib/boilerpipe/extractors/num_words_rules_extractor.rb
boilerpipe-ruby-0.4.4 lib/boilerpipe/extractors/num_words_rules_extractor.rb
boilerpipe-ruby-0.4.3 lib/boilerpipe/extractors/num_words_rules_extractor.rb
boilerpipe-ruby-0.4.2 lib/boilerpipe/extractors/num_words_rules_extractor.rb
boilerpipe-ruby-0.4.1 lib/boilerpipe/extractors/num_words_rules_extractor.rb