/trunk/inputs/VegBank/taxonobservation_/new_terms.csv - Changes - BIEN 3 - NCEAS Projects

root/trunk/inputs/VegBank/taxonobservation_/new_terms.csv

#	Date	Author	Comment
11970	01/20/2014 11:33 AM	Aaron Marcuse-Kubitza	moved everything into /trunk/ to create the standard svn layout, for use with tools that require this (eg. git-svn). IMPORTANT: do NOT do an `svn up`. instead, re-use your working copy's existing files with `svn switch` (http://svnbook.red-bean.com/en/1.6/svn.ref.svn.c.switch.html).
11524	10/31/2013 02:46 AM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/map.csv, postprocess.sql: mapped identifiedBy (the join_words() of identifiedBy_first, etc.)
11523	10/31/2013 02:34 AM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/create.sql: also join party_id to get the identifiedBy (not mapped yet). note that the inserted row count changes, because taxonobservation_ does not yet have a pkey to do a stable ordering with.
11521	10/31/2013 02:06 AM	Aaron Marcuse-Kubitza	inputs/VegBank/vegbank.~.clean_up.sql: taxoninterpretation.party_id: don't rename to taxoninterpretation_party_id, so that this can be used directly in taxonobservation_/create.sql with a USING join
11520	10/31/2013 01:52 AM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/create.sql: join taxonobservation to taxoninterpretation (as in CVS) instead of vice versa, since taxonobservation is the primary, operative table. having VegBank and CVS do things the same way helps ensure that fixes in one can transfer easily to the other.
11511	10/30/2013 09:07 PM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/map.csv: originalinterpretation, currentinterpretation: removed table name prefix so these would automap
11438	10/25/2013 09:22 AM	Aaron Marcuse-Kubitza	fix: inputs/VegBank/taxonobservation_/map.csv: remapped int_* to OMIT because these are not specific to the taxoninterpretation row (this is in a separate taxoninterpretation for the original determination instead). see wiki.vegpath.org/Spot-checking#2013-10-10 > Mike Lee's conference call feedback.
11224	10/09/2013 06:25 PM	Aaron Marcuse-Kubitza	inputs/VegBank/plantconcept_/: mapped columns, since this is now included in import_order.txt and therefore gets processed by the column-renaming runscripts. note that this means that in taxonobservation_/map.csv, the plantconcept_ input column names need to be changed to what they are mapped to.
11016	09/19/2013 11:31 AM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/map.csv: taxonomic ranks not in VegCore: removed table prefix so they will be automapped (they are globally unique)
10998	09/16/2013 07:48 AM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/create.sql: join starting with taxoninterpretation so that we can use the taxoninterpretation_id as the row_num (text strings, formed from concatenated #s cannot be used as a row_num). there is only 1 taxonobservation without a taxoninterpretation, so we can just include one row for each taxoninterpretation.
10993	09/15/2013 09:07 PM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/postprocess.sql: added primary key. note that the inserted row count changes, most likely because the rows are now in sorted order.
10944	09/12/2013 06:43 PM	Aaron Marcuse-Kubitza	inputs/VegBank/: prepended the table name to each column name to prevent column collisions, using the steps at http://wiki.vegpath.org/Left-joining_a_datasource
10931	09/12/2013 04:14 PM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/: translated multi-column filters to postprocessing derived columns, using the steps at http://wiki.vegpath.org/Adding_new-style_import_to_a_datasource#Translating-filters-to-postprocessing-derived-columns
10691	08/20/2013 10:31 AM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/map.csv: plantname: remapped to DUPLICATE#of:plantconcept_plantname because this is an exact duplicate
10690	08/20/2013 10:24 AM	Aaron Marcuse-Kubitza	bugfix: inputs/VegBank/taxonobservation_/map.csv: updated input column names for renamings in inputs/VegBank/vegbank.~.clean_up.sql
10688	08/19/2013 05:48 AM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/create.sql: also join to plantname, since plantconcept.plantname may not always be populated when plantname.plantname is
10685	08/19/2013 05:13 AM	Aaron Marcuse-Kubitza	fix: inputs/VegBank/taxonobservation_/map.csv: also mapped plantname to scientificName, since int_currplantscifull is not always provided when this is. (it cannot replace int_currplantscifull, because when int_currplantscifull also provided, this often leaves out lower ranks.) this should fill in taxonomic information for taxonobservations that are currently missing it.
10683	08/19/2013 01:53 AM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/map.csv: collector_id: remapped to UNUSED. removed LEFT JOIN collector_id->party since this field is never populated.
10680	08/18/2013 10:42 PM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/map.csv: int_origplantscifull: remapped to EQUIV (to authorplantname). this is the scrubbed originalScientificName, but we do our own scrubbing.
10678	08/18/2013 10:21 PM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/map.csv: mapped int_origplant, int_currplant to scientificName/taxonName/etc.
8236	03/28/2013 04:27 PM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/map.csv: Mapped int_currplantcommon to vernacularName
7407	01/31/2013 06:08 PM	Aaron Marcuse-Kubitza	mappings/VegCore.csv: Regenerated from wiki. This adds Brad's DwC ID terms and their definitions in <https://projects.nceas.ucsb.edu/nceas/attachments/download/621/vegbien_identifier_examples.xlsx>.
7045	01/04/2013 04:30 PM	Aaron Marcuse-Kubitza	mappings/VegCore.csv: Regenerated from wiki
6831	12/14/2012 02:46 AM	Aaron Marcuse-Kubitza	mappings/VegCore.csv: Term names: Changed special characters to _ because Redmine doesn't support special characters in HTML anchors (it removes everything except letters, numbers, _, and -)
6176	11/14/2012 05:57 PM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/map.csv: Mapped new givenname, surname (from collector_id's party) to recordedBy
5947	11/01/2012 09:29 AM	Aaron Marcuse-Kubitza	mappings/VegCore.csv: Renamed Binomial to TaxonName because this field can store more ranks than just the genus+specificEpithet binomial (that goes in speciesBinomial)
5631	10/18/2012 11:57 AM	Aaron Marcuse-Kubitza	mappings: Renamed scientificName to binomial because DwC defines the scientificName as "The full scientific name, with authorship and date information if known", but many datasources do not include the author in their scientific name, and the fields scientificName is mapped to in VegBIEN assume it does not include the author
5062	09/27/2012 08:31 AM	Aaron Marcuse-Kubitza	mappings/VegCore.csv: Renamed verbatim* taxonomic terms to original* because in most datasources, they are in fact for the original taxon determination of the organism (which can be a completely different name than the primary determination), rather than merely unscrubbed versions of the primary taxonomic name elements. Note that SALVIAS's orig_* terms do appear to be merely unscrubbed versions, but it's not a problem to add an additional taxon determination for them.
4857	09/19/2012 09:31 PM	Aaron Marcuse-Kubitza	input.Makefile: Maps validation: %/new_terms.csv: Include the entire map spreadsheet row, so that each new term is listed together with its mapping. This facilitates adding new mappings to mappings/Veg+-VegCore.csv directly from any new_terms.csv. Note that the use of `sort -u` (in lib/mappings.Makefile) causes multiline comments to be separated, leading to spurious lines for each multiline comment line.
4663	09/12/2012 05:13 PM	Aaron Marcuse-Kubitza	input.Makefile: Maps validation: $(newTerms): Fixed bug where header needed to be removed before running filter_out_ci because filter_out_ci only removes the header if it matches the vocabulary's header. Removing the header afterward can cause the first row to be removed instead if the header was already removed.
4638	09/12/2012 12:43 PM	Aaron Marcuse-Kubitza	inputs///map.csv: Changed empty mappings to self mappings, using the steps at <https://projects.nceas.ucsb.edu/nceas/projects/bien/wiki/Map_refactoring#Change-empty-mappings-to-self-mappings>. Note that in map.full.csv and VegBIEN.csv, lines that have changed are always the result of the input field's case being changed to match the case of the datasource's actual column name.
4635	09/12/2012 12:07 PM	Aaron Marcuse-Kubitza	inputs/VegBank/taxonobservation_/map.csv: Updated with new renamings of colliding join columns
4596	09/11/2012 08:22 AM	Aaron Marcuse-Kubitza	input.Makefile: Maps building: %/.map.csv.last_cleanup: $(newTerms): Remove the CSV header from the terms lists so that multiple terms lists can easily be appended together
4594	09/11/2012 08:09 AM	Aaron Marcuse-Kubitza	input.Makefile: Maps building: %/.map.csv.last_cleanup: Generate reports on new and unmapped terms in map.csv

Project

General

Profile