We're on iteration 2 for this course, and it's still in somewhat rough shape. If you plan to devote significant time to the lectures, I'd recommend waiting until next spring, when we'll be teaching iteration 3 online at http://ds-class.org.
from my perspective, a single user interface paradigm is unlikely to cover the variety of use cases that the full suite of hadoop platform services (flume, sqoop, hdfs, mapreduce, hive, pig, oozie, hbase) provides.
at cloudera, we've built and released into open source a user interface construction kit and an environment to house these various "applications". that way we can incorporate any metaphor you'd like: file browser, spreadsheet, yahoo pipes, etc.
After participating in a few of the XLDB meetups, at which the requirements for and design of SciDB were debated and discussed, I was quite excited about this project. I was kicked off of the mailing list unceremoniously once they decided they wanted to make it a business, and I haven't been allowed back on since. They're certainly abusing the term "open source" for marketing purposes here.
The reason they don't want people looking behind the curtain is that the rhetoric far exceeds the reality.
Just to be clear, I was never CTO of Facebook or anything close to it (thanks for that, though!). Dustin Moskovitz and Adam D'Angelo were the only two folks to hold that title that I know about.
And yeah, I don't want to get in trouble, but ask the right people about "Team America"...
Excel has incorporated software from Frontline Systems, rebranded as "Solver", for many years. Most heavy Excel users use it every day. I couldn't dig through the marketing speak entirely, but this product appears to be a port of Solver to C# and not open source. Of limited interest to HN readers, I'd expect.
We're especially interested in web developers who have built and deployed large, extensible applications into production environments. An interest in data visualization and analysis doesn't hurt. We also have some deep distributed storage system hacking problems.
We have a strong preference for open source experience: our team (see http://cloudera.com/about) includes core contributors from the Berkeley DB, Ganglia, Lucene/Nutch, Hadoop, and MooTools projects.
We expect you to communicate ideas clearly, exhibit preternatural intellectual curiosity across a variety of domains, write quality code, and have a consistent focus on improving yourself and the team around you.
If you're interested, drop your CV and a cover letter to [email protected].
We're on iteration 2 for this course, and it's still in somewhat rough shape. If you plan to devote significant time to the lectures, I'd recommend waiting until next spring, when we'll be teaching iteration 3 online at http://ds-class.org.
Later, Jeff