I'm working on a AWS stack based web crawler. It uses S3 and SimpleDB to coordinate instances. It's implemented in Python/Twisted, and extensible via plugins.
It has a low cost of entry in terms of code and costs and scales horizontally, so it's useful for all sorts of aggregation.
It has a low cost of entry in terms of code and costs and scales horizontally, so it's useful for all sorts of aggregation.