Versions Compared

Key

  • This line was added.
  • This line was removed.
  • Formatting was changed.
Comment: Migrated to Confluence 5.3

...

  • Initial email sent 7/29/2014 5:30:05 PM to Czarek. No written confirmation back Czarek, as of 8/25/14 (Monday).
    • Q: Copy sent to Gia? Format OK? (Unreadable format within Remedy- work with plain-text email readers?)

Even after group has confirmed ChemIT's initial suggestions (which still needs to happen), iterations and subsequent clarifications may also be required. Thus, we will need the group's continued engagement to ensure ChemIT is doing things to best meet the group's needs.

...

On external 12 TB Synology system:

  • /home (3TB total3-3.5TB total (ensure under 4TB!); on Synology)
    • Can be seen by compute nodes. Data
      • Scratch data is to be stored on compute nodes until
      work is done, then moved to /home
      • calculation is done. If not automatically deleted (aborted job, etc.) it is the researcher's responsiblitliy to delete the scratch data.
      • When calculation is complete on a computer node, than the data is written to the /home (or /notbackedup, if more space is needed temporarily) partition.
    • Data on /home must be removed when not actively needed for their a researcher's cluster work. Can Researchers can move the data to /storage, for safe-keeping and convenient access.
    • Data is backed up via EZ-Backup.
    • A Group to decide on quota for any new user gets 50GB, by default. The process to request more space is through the group's designate.
  • /storage (8TB total)
    • Can NOT be seen by compute nodes.
    • This is to store data related only to Scheraga research. No personal files.
    • Data is backed up via EZ-Backup.
    • The process to request more space is through group designate.
  • TBD (1TB total)
    • To be used for either expanding /home or /storage at a later date (when we know which needs more space sooner than the other).

...

  • /notbackedup (4TB total)
    • Can be seen by compute nodes.
    • For when a researcher needs more temporary space than an individual compute node can provide's /home directory can provide.
      • No need to expand a user's /home directory to meet their temporary "peak" needs.
    • Researcher must move relevant data off as soon as they are done with it.
    • This data is NOT BACKED UP. Use only temporarily, and only when a researcher needs "surge" space.
    • No quotas. Thus, potential for researchers to "step on each other".
  • Q: Make more robust by adding hard-ware RAID and a 2nd identical drive?
    • Cost of RAID card:
    • Cost of 2nd identical hard drive:

...

Drive is backed up via EZ-Backup, and contains these partitions:

  • OS (16 -17 GB)
    • YUM-installed applications
    • And other select other cluster-specific applications
  • swap (32 GB)
  • /data (200+GB; the rest!)
    • Applications for researchers (/software -> /data/software)
    • Software source files.

...

  • Create empty /home directories for each user. Leave emptyChemIT leaves it empty, for the researcher to purposefully populate.
  • ChemIT copies all a user's data in old Matrix to the new Matrix's /storage partition. Each researcher selectively and deliberately moves what is required for their current cluster jobs, to their new Matrix home directory.
  • OUTCOMES
    • Clean /home directories, making restores and debugging much easier over time.
    • Pawel supports this approach. He estimates perhaps 1/2 - 1 hour work for each researcher.
  • To do: Figure out dates and other timing, including complete downtime.

...

1) Default quotas for new researchers.

  • Example:
/home/storage

5GB?

50GB?

100GB

How does a user request a bigger quota?

...