buhrmann Goto Github PK
Name: Thomas Buhrmann
Type: User
Company: @graphext
Twitter: tom_gxt
Location: Madrid, Spain
Name: Thomas Buhrmann
Type: User
Company: @graphext
Twitter: tom_gxt
Location: Madrid, Spain
Hive UDF's for the data warehouse
Personal blog
Example application for analyzing Twitter data using CDH - Flume, Oozie, Hive
Cinder is a community-developed, free and open source library for professional-quality creative coding in C++.
A web application that imports Nike+ running data into its own nosql database and visualises it using interactive d3 charts.
A visual in-browser exploration tool for the connectome of the nematode worm C. elegans.
Python tools for geographic data
R interface to Nike+ API and basic statistics/visualization
scikit-learn: machine learning in Python
Twitter Sentiment Analysis
The TweetLID shared task consists in identifying the language or languages in which tweets are written. Focusing on events, and news in the Iberian Peninsula, the main focus of the task is the identification of tweets written in the 5 top languages from the Peninsula (Basque, Catalan, Galician, Spanish, and Portuguese), and English We will provide the participants of the task with a training corpus that includes approximately 15,000 tweets manually annotated with the language(s). The participants will have a month to develop and tweak their language identification systems from this training corpus. They will have apply their system on the test set afterwards, and submit the output of the system, which will be evaluated and compared to the other participants’ systems. It is worth noting that some tweets are written in more than one language (e.g., partly in Portuguese, and partly in Galician), and that the language cannot be determined in some cases (e.g., “jajaja”). The corpus also takes into account these specific cases, providing annotations such as “ca+es” (written in Catalan and Spanish), “ca/es” (it can be either Catalan or Spanish, it does not make a difference in this case), “other” (it is written in a language that is not considered in the task), o “und” (when it cannot be determined).
💫 Industrial-strength Natural Language Processing (NLP) with Python and Cython
simple statistics from the command line
Social network analysis of twitter data with hadoop, flume, hive and R (igraph).
A declarative, efficient, and flexible JavaScript library for building user interfaces.
🖖 Vue.js is a progressive, incrementally-adoptable JavaScript framework for building UI on the web.
TypeScript is a superset of JavaScript that compiles to clean JavaScript output.
An Open Source Machine Learning Framework for Everyone
The Web framework for perfectionists with deadlines.
A PHP framework for web artisans
Bring data to life with SVG, Canvas and HTML. 📊📈🎉
JavaScript (JS) is a lightweight interpreted programming language with first-class functions.
Some thing interesting about web. New door for the world.
A server is a program made to process requests and deliver data to clients.
Machine learning is a way of modeling and interpreting data that allows a piece of software to respond intelligently.
Some thing interesting about visualization, use data art
Some thing interesting about game, make everyone happy.
We are working to build community through open source technology. NB: members must have two-factor auth.
Open source projects and samples from Microsoft.
Google ❤️ Open Source for everyone.
Alibaba Open Source for everyone
Data-Driven Documents codes.
China tencent open source team.