Current news...


Mid-August update:

While the thoughts of almost everyone else turn to annual holidays, staycations, BBQs, painting the outside of the house or simply lazing in the garden the summer holiday season is actually the busiest time of year for research IT, when we get ready for the next academic year beginning in October. Apart from the summer MSc project season, usage of IT facilities is typically low at this time so it's a good time to undertake updates and upgrades; here are some of the things to look forward to in the next few months:

  • NextGen cluster upgrade: two major upgrades are planned for this cluster:

    • replacement of macomps09-16: following the upgrade of the eight macomp01-08 nodes with Dell R630 servers last autumn, another eight macomp nodes macomp09-16 have been replaced with identical R630 servers. Fitted with two 10 core CPUs and 192GB of memory each, these have added a significant boost to the cluster's job handling capacity and throughput. At present there are no scratch disks in the new nodes owing to the supply chain problems affecting the IT industry as a whole, but these will be added later when we can get hold of them. The old R410 nodes from this cluster have been moved to Huxley and will join the Stats section's Hadoop cluster which is also being rebuilt & modernised this summer (see below).

    • replacement of mablads01-16: by today's standards the 16 blade nodes mablad01-mablad16 have a very modest specification; dating from 2008, they have dual quad-core Xeon CPUs and 16 GB of memory each (a few blades that were subsequently fitted as replacments for failed blades have 32GB of installed memory). The CPUs do not support the more recent instruction sets including AVX, AVX2, etc which sometimes means compiling different versions of specialist applications and libraries specifically to run on these older processors. All 16 blades and the chassis into which they plug into will soon be replaced by 10 Dell R630 servers the same as in the macomp node upgrade (above).

  • Bazooka Hadoop cluster upgrade/rebuild: this 15 node specialist cluster began life in late 2013 running the then brand-new MapR Hadoop distribution based on Apache's software suite of the same name. Using the free community edition of MapR, this was used both for research and Big Data/AI teaching courses. Unfortunately, MapR started getting into financial difficulties in 2016 and we have had to keep both the Linux and the MapR versions frozen as of summer 2016 since no further community edition updates were released and the cluster was required every spring for a Big Data course involving many students. (Eventually MapR went bankrupt, was bought out by HP Enterprise and relaunched as an expensive commercial suite, running on HPE servers and requiring costly support contracts for continued use). In the meantime, although still suitable for teaching existing MapReduce and AI courses, the usability of the cluster for research has declined over the years since it is python2-based and does not support modern python3 applications and libraries.

  • With this year's Big Data course now completed, the decision has been taken to completely rebuild the cluster using the same hardware but with the very latest Linux and Apache Hadoop components, following a successful pilot earlier this summer using a single node psuedo-cluster server (aphrodite.ma) for teaching AI courses. Using only open source code compiled locally on the cluster ensures we are not trapped into using commercial Hadoop distributions that subsequently run into commercial problems, acquisitions, etc. At the same time the opportunity will be taken to expand the cluster by adding the eight ex-NextGen R410 servers to it (see above).

  • Stats HPC cluster upgrade: since its introduction in January this year the Stats HPC has been little-used so only the submission node fallas, with 24 processors, has remained in operation throughout with the 8 compute nodes stats01-08 being powered off. In early August it was fully powered up for an upgrade & update before the compute nodes were again shut down as a precaution during the heatwave. However, if the demand arises it is easy to power one or more nodes back on to povide the full 88 processor capability.

  • rizzuto storage upgrade: this compute server, a sister to gehrig, now has the same fast 5 disk XFS-based pool as its sibling to increase the available local storage to nearly 8 TB.

  • Huxley server room: it is now becoming clear that simple expansion of the Huxley 616 server room will not be a good solution long-term and we are now looking at relocating the entire facility to another much larger room.

Older news items:

July 5th: Early July update, RStudio upgrade for apollo Stats MSc server
June 15th: June update, Huxley server room expansion, Huxley 616 network management improvements
May 4th: April update, tape backups, development of a new Hadoop/Spark platform completed
March 16th: March update, job management for the forrest GPU server, development of a new Hadoop/Spark platform
December 23rd: a mini-HPC for Stats
August 14th: two more compute servers added to the Stats MSc compute pool
July 19th: July update: extensive updates to the Stats MSc compute servers, CUDA and cuDNN updates for nvidia4
June 19th: June update: new archival server, storage upgrade for Keaveny cluster, Firedrake build servers & Big Blue Button server
May 21st: Degond Cluster becomes a full HPC
March 24th: Degond Cluster fully upgraded
February 25th: a GPU server for the StatML CDT
February 1st: Bazooka Hadoop cluster control network switch replaced
September 24th: all Stats' general purpose compute systems now running Ubuntu 18.04
August 9th: new GPU server nvidia4 introduced
April 7th: Magma software upgraded to version 2.25-4
March 17th: a new 8 card GPU server installed
March 2nd: another backup server added
February 28th: internal network expansion, ma-backup4 and GPU servers coming soon
February 6th: more large compute servers for Stats
November 1st: NextGen is shutting down on November 1st in readiness for relocation
September 29th: matlab2018 queue has now been discontinued on the NextGen cluster
August 24th: matlab2018 queue to be discontinued on the NextGen cluster
August 14th: more servers for Stats, remote monitoring upgrades, better system status reporting and more power supplies
June 23rd: more servers in the server room, expanded Bazooka Hadoop cluster now available for use
May 29th: R Shiny server memory replaced & remote management added
April 12th: NextGen cluster Maple 2019 upgrade completed
March 16th: Planned Bazooka Hadoop cluster upgrade, reorganisation of backup servers
February 19th: ma-offsite2 now online
January 19th: Matlab 2018b upgrade ongoing
December 14th: Matlab upgrade to version R2018b started, Stats section compute & storage enhancements completed, silos3 and 4 introduced
September 18th: more local storage for Stats modal server and new PostgreSQL database server launched
August 29th: new 'du' command options, cluster R upgrade and ma-backup3
July 2nd: nvidia3 now has two GPU cards
May 15th: Early summer update
March 29th: Easter update
March 24th: spring update
March 10th: late winter update
December 15th: pre-Christmas update
November 22nd: late November update
October 8th: start of 2017/2018 academic year update
2017: Midsummer's Day update
June 16th, 2017: mid-June update
June 2nd, 2017: Early summer update
April 20th, 2017: Spring update 2
March 22nd, 2017: Early spring update
March 10th, 2017: Winter update 2
February 22nd, 2017: Winter update
November 2nd, 2016: Autumn update
October 21st, 2016: Late summer update 2
October 14th, 2016: Late summer update
February 19th, 2016: Winter update
December 11th, 2015: Autumn update
September 14th, 2015: Late summer update 2
May 2nd, 2015: Spring update 2
April 26th, 2015: Spring update
November 11th, 2014: Autumn update
September 17th, 2014: Summer update 2
July 17th, 2014: Summer update
March 15th, 2014: Spring update
November 2nd, 2013: Summer update
May 24th, 2013: Spring update
January 23rd, 2013: Happy New Year!
November 22nd, 2012: No news is good news...
November 17th, 2011: A revamp for the Maths SSH gateways
September 7th, 2011: Failed systems under repair
August 14th, 2011: Introducing calculus, a new NFS home directory server for research users
July 19th, 2011: a new staging server for the compute cluster
July 19th, 2011: A new Matlab queue and improved queue documentation
June 30th, 2011: Updated laptop backup scripts
June 18th, 2011: More storage for the silo...
June 16th, 2011: Yet more storage for the SCAN...
June 10th, 2011: 3 new nodes added to the Maths compute cluster
May 21st, 2011: Announcing SCAN large storage and subversion (SVN) servers
May 26th, 2011: Reporting missing scratch disk on macomp01
May 21st, 2011: Announcing completion of silo upgrades
May 16th, 2011: Announcing upgrades for silo
April 14th, 2011: Goodbye SCAN 3, hello SCAN 4
March 26th, 2011: quickstart guide to using the Torque/Maui cluster job queueing system
March 9th, 2011: automatic laptop backup/sync service, new collaboration systems launched
May 20th, 2010: Scratch disks are now available on all macomp and mablad compute cluster systems
March 11th, 2010: Introduing job queueing on the Fünf Gruppe compute cluster
October 16th, 2008: Introduing the Fünf Gruppe compute cluster
June 18th, 2008: German City compute farm now expanded to 22 machines
February 7th, 2008: new applications on the Linux apps server, unclutter your desktop
November 13th, 2007: aragon and cathedral now general access computers, networked Linux Matlab installation upgraded to R2007a
September 14th, 2007: Problems with sending outgoing mail for UNIX & Linux users
July 23rd, 2007: SCAN available full-time over the summer vacation, closure of Imperial's Usenet news server
May 15th, 2007: Temporary SCAN suspension, closure of the Maths Physics computer room, new research computing facilities
January 14th, 2005: Exchange mail server upgrade, spam filtering with pine and various other enhancements


Andy Thomas

Research Computing Manager,
Department of Mathematics

last updated: 13.8.2022