Skip to content

Miscellaneous tools for processing WARC files from the CommonCrawl

License

Notifications You must be signed in to change notification settings

JudithBarnett/warc-tools

 
 

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 

History

11 Commits
 
 
 
 
 
 
 
 
 
 

Repository files navigation

Warc Tools

Some rather use-case-specific tools for pulling stuff out of the Common Crawl data on AWS.

License

This code is Licensed under the MIT License

Copyright © 2013 Kevin Bullaughey

About

Miscellaneous tools for processing WARC files from the CommonCrawl

Resources

License

Stars

Watchers

Forks

Releases

No releases published

Packages

No packages published

Languages

  • Go 100.0%