In this episode I explain how a community detection algorithm known as Markov clustering can be constructed by combining simple concepts like random walks, graphs, similarity matrix. Moreover, I highlight how one can build a similarity graph and then run a community detection algorithm on such graph to find clusters in tabular data.
You can find a simple hands-on code snippet to play with on the Amethix Blog
Enjoy the show!
References
[1] S. Fortunato, “Community detection in graphs”, Physics Reports, volume 486, issues 3-5, pages 75-174, February 2010.
[2] Z. Yang, et al., “A Comparative Analysis of Community Detection Algorithms on Artificial Networks”, Scientific Reports volume 6, Article number: 30750 (2016)
[3] S. Dongen, “A cluster algorithm for graphs”, Technical Report, CWI (Centre for Mathematics and Computer Science) Amsterdam, The Netherlands, 2000.
[4] A. J. Enright, et al., “An efficient algorithm for large-scale detection of protein families”, Nucleic Acids Research, volume 30, issue 7, pages 1575-1584, 2002.
Apache Arrow, Ballista and Big Data in Rust with Andy Grove RB (Ep. 160)
GitHub Copilot: yay or nay? (Ep. 159)
Pandas vs Rust [RB] (Ep. 158)
A simple trick for very unbalanced data (Ep. 157)
Time to take your data back with Tapmydata (Ep. 156)
True Machine Intelligence just like the human brain (Ep. 155)
Delivering unstoppable data with Streamr (Ep. 154)
MLOps: the good, the bad and the ugly (Ep. 153)
MLOps: what is and why it is important Part 2 (Ep. 152)
MLOps: what is and why it is important (Ep. 151)
Can I get paid for my data? With Mike Andi from Mytiki (Ep. 150)
Building high-growth data businesses with Lillian Pierson (Ep. 149)
Learning and training in AI times (Ep. 148)
You are the product [RB] (Ep. 147)
Polars: the fastest dataframe crate in Rust - with Ritchie Vink (Ep. 146)
Apache Arrow, Ballista and Big Data in Rust with Andy Grove (Ep. 145)
Pandas vs Rust (Ep. 144)
Concurrent is not parallel - Part 2 (Ep. 143)
Concurrent is not parallel - Part 1 (Ep. 142)
Backend technologies for machine learning in production (Ep. 141)
Create your
podcast in
minutes
It is Free
Insight Story: Tech Trends Unpacked
Zero-Shot
Fast Forward by Tomorrow Unlocked: Tech past, tech future
The Unbelivable Truth - Series 1 - 26 including specials and pilot
Elliot in the Morning