Big Data for Beginners: Tutorial Series
Covers data warehouse concepts; beginner tutorials for Hadoop, ZooKeeper, Hive, Flume, Kafka, Hbase, Sqoop, Oozie, Azkaban, Kylin, CDH, Impala, Hue, ClickHouse, Kettle, Ambari, ELK, Scala and Flink.
219 posts
Covers data warehouse concepts; beginner tutorials for Hadoop, ZooKeeper, Hive, Flume, Kafka, Hbase, Sqoop, Oozie, Azkaban, Kylin, CDH, Impala, Hue, ClickHouse, Kettle, Ambari, ELK, Scala and Flink.

An IPA signed with an iOS enterprise certificate can't be listed on the App Store. It can only be distributed out-of-band — for instance, by putting a QR code on your company website that Apple users scan to install. Between the IPA file and the QR code you need a plist file describing your app, and installation goes through the itms-services protocol. That's what this tutorial covers.

At the end of 2020 CentOS announced it would end support at the end of 2021, which sent shockwaves everywhere. Developers and users alike rejected it — even CentOS's founder started the Rocky Linux project that same month to build the next CentOS. Red Hat immediately adjusted course, publishing a post on its official blog on January 20, 2021.

On January 5, 2021 GitHub's official blog published "Advancing developer freedom: GitHub is fully available in Iran". Iranian developers had previously been blocked because of US sanctions.

You can't talk about data warehousing without talking about ETL. The term is most common in warehouses, but it isn't limited to them. It's a critical step, so here's a quick introduction to how ETL extraction and comparison work.

I connected my Spring Boot main site with a Discuz! forum so login and logout are synchronized, and member data from Discuz! can be shown back on my site. I also modified the previously open-sourced discuz-ucenter-api-for-java to fit the way Spring Boot is written today.

The last post covered Spring Boot back-end optimization and technology choices. This one is about front-end page performance optimization, mainly targeting the criteria Google publishes. Baidu's white paper isn't included.

The last post covered buying and configuring the server. This one is about Spring Boot optimization and technology choices: I switched to the more efficient FreeMarker, added Redis caching, added performance statistics, and added GitHub Actions workflows.

After upgrading to a newer Druid, the logs kept throwing errors: discard long time none received connection., jdbcUrl: blah blah. The program ran fine, but staring at a wall of errors is unbearable, so I went into their source to see what was actually going on.

Regarding this outage, some netizens claimed that, due to the pandemic, servers in various regions had been stolen. However, this claim has now been debunked — it was just a Photoshopped image.
