Language Selection

English French German Italian Portuguese Spanish

Mining DistroWatch.com Logs (Part 2)

Filed under
Linux

This article pursues the analysis of DistroWatch.com's logs I started one week ago. Last time, the data were prepared so that we could investigate the evolution, in time and space, of the popularity of GNU/Linux distributions. Pre-processing the logs in a different manner allows to focus on other interesting questions. In this way, although the extracted patterns will have the same "shape" as in last week's extraction, they will, this time, help us in discovering groups of distributions fulfilling similar purposes.

Instead of last week's ternary relation, this time, we will end up with mining a 4-ary relation. More precisely, a symmetric graph of distributions evolving in time and space. Take the red pill and welcome to the real world... of data-mining! When a visitor of DistroWatch.com (identified by her IP address from which the country is inferred) visits, the same day, pages related to different distributions, she probably searches for a Free operating system to fulfill her specific needs.

Among the millions of visits in a semester, almost all pairs of distribution pages have, one day, been consulted by at least one visitor.

Let us start with the biggest community: the old mainstream general-purpose distributions. At the center of this community (again remember that this does not relate to popularity but to the common purposes these distributions serve), Slackware, Gentoo and Ubuntu. A bit further away (i.e., not as much related to the other distributions of this group), Fedora, openSUSE and Debian. At the border of this community, Yellow Dog, MEPIS, Mandriva, Vector, FreeBSD and Damn Small Linux. When looking at the countries present in these patterns, it appears that the visitors from some European countries are clearly those making these associations. The United Kingdom shows off by being in almost all these patterns. Finland is also extremely present. Australia, Greece and Denmark are not far away. Why would these European and Australian visitors focus more on mainstream distributions than others? Maybe they are more conservative and keep on tracking the evolution of these solid distributions instead of searching for more specialized ones.

More Here




More in Tux Machines

digiKam 7.7.0 is released

After three months of active maintenance and another bug triage, the digiKam team is proud to present version 7.7.0 of its open source digital photo manager. See below the list of most important features coming with this release. Read more

Dilution and Misuse of the "Linux" Brand

Samsung, Red Hat to Work on Linux Drivers for Future Tech

The metaverse is expected to uproot system design as we know it, and Samsung is one of many hardware vendors re-imagining data center infrastructure in preparation for a parallel 3D world. Samsung is working on new memory technologies that provide faster bandwidth inside hardware for data to travel between CPUs, storage and other computing resources. The company also announced it was partnering with Red Hat to ensure these technologies have Linux compatibility. Read more

today's howtos

  • How to install go1.19beta on Ubuntu 22.04 – NextGenTips

    In this tutorial, we are going to explore how to install go on Ubuntu 22.04 Golang is an open-source programming language that is easy to learn and use. It is built-in concurrency and has a robust standard library. It is reliable, builds fast, and efficient software that scales fast. Its concurrency mechanisms make it easy to write programs that get the most out of multicore and networked machines, while its novel-type systems enable flexible and modular program constructions. Go compiles quickly to machine code and has the convenience of garbage collection and the power of run-time reflection. In this guide, we are going to learn how to install golang 1.19beta on Ubuntu 22.04. Go 1.19beta1 is not yet released. There is so much work in progress with all the documentation.

  • molecule test: failed to connect to bus in systemd container - openQA bites

    Ansible Molecule is a project to help you test your ansible roles. I’m using molecule for automatically testing the ansible roles of geekoops.

  • How To Install MongoDB on AlmaLinux 9 - idroot

    In this tutorial, we will show you how to install MongoDB on AlmaLinux 9. For those of you who didn’t know, MongoDB is a high-performance, highly scalable document-oriented NoSQL database. Unlike in SQL databases where data is stored in rows and columns inside tables, in MongoDB, data is structured in JSON-like format inside records which are referred to as documents. The open-source attribute of MongoDB as a database software makes it an ideal candidate for almost any database-related project. This article assumes you have at least basic knowledge of Linux, know how to use the shell, and most importantly, you host your site on your own VPS. The installation is quite simple and assumes you are running in the root account, if not you may need to add ‘sudo‘ to the commands to get root privileges. I will show you the step-by-step installation of the MongoDB NoSQL database on AlmaLinux 9. You can follow the same instructions for CentOS and Rocky Linux.

  • An introduction (and how-to) to Plugin Loader for the Steam Deck. - Invidious
  • Self-host a Ghost Blog With Traefik

    Ghost is a very popular open-source content management system. Started as an alternative to WordPress and it went on to become an alternative to Substack by focusing on membership and newsletter. The creators of Ghost offer managed Pro hosting but it may not fit everyone's budget. Alternatively, you can self-host it on your own cloud servers. On Linux handbook, we already have a guide on deploying Ghost with Docker in a reverse proxy setup. Instead of Ngnix reverse proxy, you can also use another software called Traefik with Docker. It is a popular open-source cloud-native application proxy, API Gateway, Edge-router, and more. I use Traefik to secure my websites using an SSL certificate obtained from Let's Encrypt. Once deployed, Traefik can automatically manage your certificates and their renewals. In this tutorial, I'll share the necessary steps for deploying a Ghost blog with Docker and Traefik.