Commit Graph
24 Commits
Author SHA1 Message Date
Greg Von Kuster bac7fc59be New script to rename purged datasets. 2008-02-02 01:59:44 +00:00
Greg Von Kuster cdd0f9c1e1 Next pass at cleaning up datasets and histories.
- Added an option to remove dataset from disk when purging, or leave them for renaming and later removal
- Moved History.purge() and Dataset.purge() methods to the cleanup_datasets.py script
2008-02-01 20:07:11 +00:00
Greg Von Kuster 8d07f84003 Another pass at the cleanup datasets stuff.
- Eliminated duplicate functionality from admin controller, moving appropriate components to Galaxy reports.
- Re-implemented the long-term process we will use for deleting / purging histories and datasets
  (we will initially write external shell scripts to rename dataset files rather then deleting them)
2008-01-30 20:12:43 +00:00
Greg Von Kuster c8ee253ba6 Missed committing similar fixes to cleanup_datasets script and admin controller. 2008-01-27 16:41:29 +00:00
Nate Coraor 7d8656b55c First draft of the fetch_eggs.py script, this will download eggs from a
distribution point and copy them to Galaxy's eggs/ directory.

fetch_eggs_tarball.py and check_python_platform.py are mostly for offline
systems.  fetch_eggs_tarball.py fetches eggs for your platform (or a
specified platform) and then tars them up, for copying to another system.
check_python_platform.py returns the string for your platform/interpreter
that will be used to download from the egg distribution site.
2008-01-25 18:39:00 +00:00
Greg Von Kuster be461b0c8a Eliminated "removed" state from dataset lifecycle, require db schema update for those that installed rev 2256.
For those that may have installed revision 2256, use the following SQL command to evolve your schema:
alter table dataset drop column removed;

We have changed direction again on the process we will use to cleanup datasets and histories.
- When we are comfortable with the purge process, the Dataset.purge() function will be modified to delete the file from disk rather than simply renaming it, and the Dataset.remove_from_disk() function will be eliminated.
- The entire process below must be executed at least 1 time ( prior to making these changes ) to ensure that purged files are properly removed from disk.  This process is:

delete_userless_histories() -> purge_histories() -> purge_datasets() -> remove_datasets()

Shell scripts have been added for each of these stages with the exception of remove_datasets(), which will eventually be eliminated.
2008-01-24 13:56:09 +00:00
Greg Von Kuster 5ea5d75da5 Enhancements to history and dataset cleanup (requires db schema change).
SQL commands to alter schema:
alter table dataset add column removed boolean default 'f';
alter history add column purged boolean default 'f';

- The lifecycle of a history after creation is now deleted -> purged
- Deleting a history will delete all associated datasets (only the database is update, nothing removed from disk).
- Purging a history will purge all associated datasets (datasets are renamed, nothing removed from disk).
- The lifecycle of a dataset after creation is now deleted -> purged -> removed
- Deleting a dataset will only update the database.
- Purging a dataset will rename the file on disk and update the database.
- Removing a dataset will remove the renamed, purged file from disk and update the database.
- Also reverted the behavior of history_delete() to pre-rev 2250 so that no datasets are deleted at teh time the history was deleted by the user.
- The admin controller and the cleanup_datasets.py script include all of the above functionality.
2008-01-22 20:47:04 +00:00
Greg Von Kuster a2b76f55f5 Purging datasets will now only rename them ( by appending the string '_purged' to the original file name ) rather than removing them. Purged dataset files can then be manually deleted from the file system at a later time. 2008-01-21 16:11:12 +00:00
Greg Von Kuster 0a1bdbffe9 Added 2 new options to cleanup_datasets that provide info about what will be abandoned or purged.
Also now using Dataset.create_time in the purge() function since Dataset.update_time would have been reset when the abandon() function was executed.
2008-01-17 19:40:51 +00:00
Greg Von Kuster 2dc1c17a2d Cleaned up code for cleaning up histories and datasets, should now be faster and require much less memory. Other miscellanwoud code cleanup. Added sample database config options to config file. 2008-01-16 19:37:30 +00:00
Greg Von Kuster 01378424b4 New reports app server and 1st report showing disk usage for dataset file system - still need to mess with styles a bit.
Also fixed update_dataset_size script and delete deprecatd random_intervals code file.
2008-01-11 21:24:21 +00:00
Greg Von Kuster 6d80665988 Added file_size column to dataset table, requires db schema alteration:
alter table dataset add column file_size numeric(15,0)
Also added new script update_dataset_size.py
2008-01-08 16:19:25 +00:00
James Taylor 4b88be5e25 1) Modified index page so progressive loading works better for slower
connections

2) Added a script for packing the javascripts, everything is now repacked
   with yui-compressor (the packed versions are slightly large, but they
   gzip better resulting in an overall reduction).
2007-12-11 19:30:03 +00:00
James Taylor a50b8007ef Splitting out UCS2 and UCS4 eggs for linux. Run scripts check
waht python is available and use the right directory. Factored
path stuff out of shell scripts into a shared 'setup_paths.sh'.
2007-12-06 01:27:33 +00:00
Daniel Blankenberg 74a9f2ecbf Modify metadata update script to update metadata of all tabular based files. Special case for interval files (interval,
bed) to not override interval specific assignments.
2007-10-17 15:23:09 +00:00
Greg Von Kuster 5aa8cb1e23 Changed the call from Tabular().set_meta() to dataset.set_meta() since we now have a better approach to setting meta data for tabular data types. 2007-09-28 18:42:57 +00:00
Daniel Blankenberg ed964dd043 Add dataset filename table. Filenames can now be assigned, setting the readonly flag on a dataset filename object
will prevent Galaxy from deleting the file when the dataset is purged.  Copies of datasets (such as when sharing a
history), no longer copy file contents.

Database changes required:

CREATE TABLE dataset_filename (
        id INTEGER NOT NULL,
        create_time TIMESTAMP DEFAULT current_timestamp,
        update_time TIMESTAMP DEFAULT current_timestamp,
        filename TEXT,
        extra_files_path TEXT,
        readonly BOOLEAN,
        PRIMARY KEY (id)
);

ALTER TABLE dataset ADD filename_id INTEGER;
2007-09-24 19:42:09 +00:00
Daniel Blankenberg 42a90ea1f0 Script to ensure that MAF metadata is set. 2007-08-30 14:46:52 +00:00
Daniel Blankenberg b873d64859 Scripts to update metadata to the latest model.
Be sure to backup databases before executing.
2007-08-22 20:04:40 +00:00
Daniel Blankenberg 86651d8a66 Add a commandline util to cleanup abandoned histories and purge deleted datasets. 2007-07-31 21:40:14 +00:00
Ian Schenck fda163bbe8 Data validation. Please please please take a look and give me some input. 2007-03-12 18:35:03 +00:00
Ian Schenck 496ba08a30 Validating on upload. Still need to make template display a peek or something. 2007-03-09 21:02:23 +00:00
Ian Schenck 2ddab5819a Script that will run like a tool for each dataset generated. Is this in the right place? P.S. this stuff isn't final. 2007-03-09 19:45:29 +00:00
James Taylor 2c4c20436e Cleaning up modules. Still a few stragglers that can't be easily eggified.
EVERYTHING we depend on should eventually be under the galaxy module or
packaged in eggs to avoid namespace problems.
2006-11-15 18:27:21 +00:00