I used AI to build a SaaS, and here’s my experience

Tech

by Sunny Srinidhi - July 16, 2025July 16, 20250

A seasoned developer reflects on his journey with coding, particularly frontend technologies. Despite initial fears surrounding frameworks like React, he successfully embraced Swift for personal projects, enhancing his coding skills. However, his experience using AI tools like Cursor for web apps revealed limitations and frustrations, leading him back to Swift for reliability and confidence.

The Impact of Generative AI on Data Engineering

Data Science

by Sunny Srinidhi - March 9, 2025March 9, 20250

Generative AI is transforming the field of data engineering by automating complex processes such as data augmentation, cleaning, integration, and anomaly detection. Unlike traditional AI, which focuses on analysis and prediction, Generative AI creates new data based on learned patterns. This capability improves data quality, enhances efficiency, and enables scalable solutions. However, challenges like data privacy, model bias, and ethical concerns must be carefully managed. As AI technology advances, its role in data engineering will continue to expand, leading to more intelligent and automated data workflows.

Data Automation with AI/ML: A Comprehensive Guide

by Sunny Srinidhi - November 28, 20240

The article discusses the transformative impact of artificial intelligence (AI) and machine learning (ML) on data automation, enhancing efficiency, decision-making, and scalability in businesses. It explores trends like generative AI, AutoML, data governance, and democratization while providing real-world applications across various industries, ultimately guiding businesses in effective AI/ML integration.

Lemmatization in Natural Language Processing (NLP) and Machine Learning

Data Science

by Sunny Srinidhi - February 26, 2020February 26, 20200

Lemmatization is one of the most common text pre-processing techniques used in Natural Language Processing (NLP) and machine learning in general. If you've already read my post about stemming of words in NLP, you'll already know that lemmatization is not that much different. Both in stemming and in lemmatization, we try to reduce a given word to its root word. The root word is called a stem in the stemming process, and it is called a lemma in the lemmatization process. But there are a few more differences to the two than that. Let's see what those are. How is Lemmatization different from Stemming In stemming, a part of the word is just chopped off at the tail end to arrive at

Stemming of words in Natural Language Processing, what is it?

Data Science

by Sunny Srinidhi - February 19, 2020August 27, 20241

Stemming is one of the most common data pre-processing operations we do in almost all Natural Language Processing (NLP) projects. If you're new to this space, it is possible that you don't exactly know what this is even though you have come across this word. You might also be confused between stemming and lemmatization, which are two similar operations. In this post, we'll see what exactly is stemming, with a few examples here and there. I hope I'll be able to explain this process in simple words for you. Stemming To put simply, stemming is the process of removing a part of a word, or reducing a word to its stem or root. This might not necessarily mean we're reducing a word

Removing stop words in Java as part of data cleaning in Artificial Intelligence

Data Science

by Sunny Srinidhi - February 5, 2020February 5, 20200

More in The fastText Series. Working with text datasets is very common in data science problems. A good example of this is sentiment analysis, where you get social network posts as data sets. Based on the content of these posts, you need to estimate the sentiment around a topic of interest. When we're working with text as the data, there are a lot of words which we want to remove from the data to "clean" it, such as normalising, removing stop words, stemming, lemmatizing, etc. In this post, we'll see how we can remove stop words from our input text to clean our data so that our analysis is based only on the actual content of the data. But wait, what are stop

An Intro to Affective Computing

Data Science

by Sunny Srinidhi - January 7, 2020January 7, 20200

Not a lot of us have heard of Affective Computing. Most people I have spoken to about this didn't know anything about Affective Computing. So I thought, I'll just write an intro, explaining what I have understood about the discipline and hopefully, will get to learn more from the comments. So let's get started. Affecting computing is all about understanding human emotions in a human-machine interface system and responding based on those emotions. Consider this, you get into an ATM vestibule to draw some cash, but you're tensed about getting late to your date, who is already waiting for you at the restaurant. If anybody sees you in this condition at the ATM vestibule, they'll be able to easily understand that

Optimising a fastText model for better accuracy

Data Science

by Sunny Srinidhi - December 3, 2019December 19, 20190

More in The fastText Series. In our previous post, we saw what n-grams are and how they are useful. Before that post, we built a simple text classifier using Facebook’s fastText library. In this post, we’ll see how we can optimise that model for better accuracy. Precision and Recall Precision and recall are two things we need to know to better understand the accuracy of our models. And these two things are not very difficult to understand. Precision is the number of correct labels that were predicted by the fastText model, and recall is the number of labels, out of the correct labels, that were successfully predicted. That might be a bit confusing, so let’s look at an example to understand it better. Suppose for a sentence

Understanding Word N-grams and N-gram Probability in Natural Language Processing

Data Science

by Sunny Srinidhi - November 26, 2019December 19, 20192

More in The fastText Series. N-gram is probably the easiest concept to understand in the whole machine learning space, I guess. An N-gram means a sequence of N words. So for example, “Medium blog” is a 2-gram (a bigram), “A Medium blog post” is a 4-gram, and “Write on Medium” is a 3-gram (trigram). Well, that wasn’t very interesting or exciting. True, but we still have to look at the probability used with n-grams, which is quite interesting. Why N-gram though? Before we move on to the probability stuff, let’s answer this question first. Why is it that we need to learn n-gram and the related probability? Well, in Natural Language Processing, or NLP for short, n-grams are used for a variety of things.

An intro to text classification with Facebook’s fastText (Natural Language Processing)

Data Science

by Sunny Srinidhi - November 25, 2019December 19, 20193

More in The fastText Series. Text classification is a pretty common application of machine learning. In such an application, machine learning is used to categorise a piece of text into two or more categories. There are both supervised and unsupervised learning models for text classification. In this post, we’ll see how we can use Facebook’s fastText library for some simple text classification. fastText, developed by Facebook, is a popular library for text classification. The library is an open source project on GitHub, and is pretty active. The library also provides pre-built models for text classification, both supervised and unsupervised. In this post, we’ll check out how we can train the supervised model in the library for some quick text classification. The library