Home / Companies / DataStax / Blog / Post Details
Content Deep Dive

When Rotten Tomatoes Isn't Enough: Twitter Sentiment Analysis with DSE Part 3

Blog post from DataStax

Post Details
Company
Date Published
Author
Amanda Moran
Word Count
2,401
Company Posts That Month
15
Language
English
Hacker News Points
-
Post removed?
No
Summary

Part 3 of the blog series explores the use of DataStax Enterprise Analytics and various technologies like Apache Cassandra, Apache Spark, PySpark, Python, and Jupyter Notebooks for conducting text analytics, specifically sentiment analysis on movie-related Twitter data. The process involves setting up the necessary environment, including DataStax and the Twitter Developer API, then using Apache Spark’s MLlib functions such as Tokenizer and StopWordsRemover to preprocess the tweets. The analysis is performed using the Pattern library in Python to determine sentiment scores from cleaned tweets, which are stored in Cassandra tables. By comparing positive and negative sentiment scores, the analysis concludes whether a movie is generally liked based on the Twitter data, exemplified by an analysis of "SpiderVerse," which was found to be positively rated. The blog emphasizes the iterative nature of data science and encourages further exploration and feedback.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Serverless 1 344 44 29 -37%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.