Blog post

Showing posts with label source based data entry. Show all posts
Showing posts with label source based data entry. Show all posts

Friday, January 1, 2010

Software review - Geves

I've only found two software packages that attempt the kind of source-based data entry I describe in my previous post (not counting products such as Custodian or Clooz, as their purpose is different). They are 'Genealogy Research System' and another that I found a few weeks ago, 'Geves'. I've been having a ball playing with Geves since I downloaded the trial. I used the deluxe version, which has more customisation options that the standard version. The version I downloaded was 1.3.28.

Just quickly before I begin - I have no financial interest in this product other than as a customer and have not been offered anything by anyone to write this. I'm writing the review purely because I found the program interesting. 

Geves - how it works 

The good

While it's possible to go to an individual person's view and enter information, Geves is based around the idea of entering data from a source document. The best way to start is not adding details to a person view, but by choosing the type of source you want to insert.





The program gives a reasonable selection of data entry forms for England, Scotland and Wales and some general purpose forms. If the birth death or marriage form you want isn't there (eg I need Australian certificates) you can choose the most similar form and adapt it to your needs. That didn't work too badly with my Australian BMD records for data entry but they do end up labelled as the wrong type of certificate which is annoying.

Here's an almost completed form for a Scottish ancestor, James Couper, who was kind enough to pass away in 1855 - giving me a death certificate with more than usual detail! If you attach an image to the source record, Geves displays it in a split screen below the data entry form. It makes data entry and checking against the source very easy.


The yellow dropdown boxes are where you choose which individual or place the name in the record relates to - or you can add a new person at a click of a button. If you do choose from the dropdown, a box appears when you hover over a name with the events from that person's life. You can also use the extensive list and search windows to find the correct person in the database, then drag and drop their name across to the box.

I very much like the adjustable split screen source display. The forms aren't too bad, so far as they go. I did find occasionally that some detail seemed to be missing but that's relatively easy to correct. There is some data validation, with cells coming up in shades of orange or red if you accidentally try to make someone their own grandfather. The data entry was quick and easy - completing and checking the form took only a few minutes.

Individual and place names are as recorded on the source. You make the link to a individual in the database or a place on the map for the purpose of searches. So far so good, and it's all quite efficient. But... it gets better.

The very good
The program includes a web browser. For certain records (primarily UK census) on selected genealogy sites you can download the information into the appropriate form at the click of a button, including the source image(s). It takes about 20 seconds and all that's left to do is correct the information against the image (add in any data not in the transcription), and link it up to the right people and places. Even a fairly large household took me less than five minutes to enter this way, from start to finish. I was mostly using FindMyPast. When I tried Ancestry, Geves managed to suck information into the form that wasn't visible on the page. Clever.

In less than two weeks, when I was looking after a sick baby and was also unwell myself, I still managed to enter every detail for 125 UK census households without breaking a sweat. 125. Every detail. Really! I told you I had a ball with it! I got a kick out of seeing the forms fill themselves out every time.

The other very nice thing is that if you do enter a source only to find that the John Smith in that source isn't the John Smith you thought, it's very easy to undo. Just clear the dropdown box where you identify who (in your database) the John Smith in the source relates to. All the other bits and pieces from the source - residence, employment, birth year, whatever - will disappear from your John Smith's record.

You can't please all the people all the time...
As you enter data into the source template Geves translates it into events for you, including lineage connections between the people. Obviously the program has to make a lot of assumptions along the way, and they may not always be assumptions you agree with. If this only happens occasionally it's no big deal to fix. On the other hand all the time and effort saving made in data entry would be lost if you disagreed with something the program did in a systematic way, as there is nowhere (that I saw) to vary any of the assumptions that Geves makes.

Here are the events automatically created by Geves for the death certificate above...



 

And here's the family tree it has constructed...

 


The program does a great job where the family connections are defined by the form (eg father and mother on a birth certificate) including spouse and parent-child relationships to the head person on a census form. It makes no attempt (and nor should it) to make more ambiguous connections eg a niece or stepson. You can quite quickly and easily add events to the source yourself. For example, a birth record for the niece with just the name of the child and the parents. You don't need to put in any details that are already generated by the form, such as estimated birthdate or place of a birth for a niece on a census form. There are a few more issues along these lines for census forms in particular, and it is worth reading the help file to be aware of them and understand how to work around them. 

Geves does not have all the event types you usually find. Census entries were recorded as a residence, or "Visited at" depending on the description. I would prefer to use a census event and say that the person was "enumerated at". Maybe that's just me. It does have a research screen to see what censuses you have found for each person so I suppose I could live with it - but there may well be other decisions that grate.

Once you have multiple sources for an event, Geves makes more assumptions about which source is more reliable in order to combine the information. Unfortunately there is little description of how this decision is made, either. It's obvious which events are merged, though, and you are able to change the outcome of the merged event or mark it as confirmed.

Here's a screenshot of the events tab for James Couper after I added 1841 and 1851 census information.




Merged events have a blue square around the icon, and are marked "Merged". As you click on events, you can see the source (or sources) for the information in that event underneath.

To be fair, making assumptions is always going to be a perilous business. Overall I think this part of the program is quite well-handled.

The not-so-good and really-quite-bad...

The first not-so-good thing that I noticed about Geves was the very basic output options. From a person page you can access a few simple lists of events (not a lot of detail), an ancestor report and a descendent report. There are no options about what should or should not be included in these. You can change the font, but that's it. You can also print out lists which are really very good, with whatever fields and sorting you want, but it's not the same thing.

Worse, though, there were no source citations on the reports! I could hardly believe that!

It occured to me that you could enter data here, then use another program to do reports and charts if only the GEDCOM export worked nicely with your other program. I shall defer here to someone with better knowledge than myself. Tamura Jones reviewed a beta copy of Geves back in 2007 and gave the program a rating of "Dismal" citing, among other things, GEDCOM import and export problems. I don't know if any of the problems mentioned have been fixed since then, but the ones I was able to check on with my definately non-expert knowledge of GEDCOM had not been. This was disappointing. Source-based data entry loses its gloss if you're not sure of getting the information out again! 

The other stuff
Otherwise, I had no particular problems using the program. It has been stable for me. There was a full 5 second delay going to and from one particular source with 24 sub-records but that was the only lag I noticed. I'm not sure how the program would handle large files. My experimental file was relatively small with just over 700 individuals (many of whom were probably the same people, but I hadn't made the connection yet) and aside from the delay I mentioned I didn't have any problem.

Although the general reports were lacking, the lists were good and very editable. Some fields were easy to add, some you had to start writing expressions. I get the impression that the lists could be very powerful, if only you could work out how to write the expressions. There was some information on that in the help file.

What else...? Photos are treated like source images. The source form relating to the photo has fields for date, place, event, photographer and allows you to mark with a box (which then becomes a passport photo for the individual) who is in the picture. If using the deluxe version you can add your own custom fields to sources, people, events, repositories or whatever. I think I would want to add a custom caption or description field to the source record. It was difficult to find how to access the custom forms you created. I eventually found the answer not in the help file but on the user forum. As an aside, judging by the forum the user base is very small, but the developer is quick to give a helpful response to questions.

I don't know how those custom fields are treated in GEDCOM export. There did seem to be an ability to customise how sources were exported for each template. I didn't experiment with it, but if using the program to export I would. The default export for the source was the one line description given to it although all the other information you would need for a proper source citiation was generally tucked away on the source forms.

The conclusion
I haven't covered every aspect of the program here. In some ways it has surprising depth and is very customisable, in others it's inflexible and falls a little short. Although I have raised concerns and criticisms of the program, I did buy a copy at the end of the free 30 day trial.

Yes, I am perhaps a little besotted with the data entry to the exclusion of all else. They do say love is blind.

I'm not planning on using Geves as my primary genealogy software. I used it in the trial period to begin sorting out the Tregoning families in Gwennap, Cornwall, England. At least 2 of the 17 households with Tregonings in 1841 are of interest to me.

Thanks to Geves I am now confident that I've followed through to the correct 1881 census entry for my great-great-great grandmother, Mary Tregoning, and I think I know which death record is hers, too. I'll enter just the information relating to my family into my main database. It's going to seem so sloooow after Geves, but at least I'll be able to assume whatever I want and make some pretty charts. With source citations.

By the way, Happy New Year!

Friday, December 18, 2009

My vision for genealogy data entry

If I had written a Genea-Santa letter, I would have asked for perfect genealogy software. One of the features of my utopian software (and there's a long list) is source-based data entry.

Here's my thinking... 
Sources are important. I don't enter information into my genealogy software unless I have a source. That source may be something as simple as a note I write to myself (eg noting a conversation with a relative) but it will be something I can use to identify where a particular piece of information came from, and how reliable that information is likely to be.

A lot of the source documentation we use come in forms of one sort or another. Births, deaths, marriages, census information - all entered into forms.

When it comes time to do data entry, there I am sitting in front of my computer with a source document in my hand or on my screen. It's quite likely that the information will be recorded in a form that I have seen before, and I will see again. Typically, I will set up a source record and lock it on while I am entering the data. Then it's a matter of going through the source document in a systematic way, navigating to the appropriate person in my database, and adding or editing events in the person's life.

The way I described it, it sounds pretty efficient. It's not really. There's an awful lot of moving from person to person, editing a bit, moving around again, finding your place on the form, not to mention re-entering the same information over and again (eg an address on a census form). Every time you move around you are distracted from where you are on the form. Every time you have to enter a piece of information again, you may enter it differently. If you want to check over your data entry you have to navigate around all over again.

What if you enter the information only to realise you had the wrong person? Who would have thought there could be more than one John Smith?! Then you have to track down and undo all those little changes you made.


It's slow. It's prone to error. It's hard to check. It's hard to undo.

A feature of my imaginary ideal genealogy software is the ability to enter data, where possible, in the template of the source document. The act of entering the data should generate all the citation details (maybe add a field or two for anything relevant not on the form itself, eg repository) and should handle the data entry. You would enter the data once.  Perhaps I'm fundamentally lazy, but if I have typed something in once, I don't want to have to type it in again.

Data entry would be quick and easy because you would not have to constantly find your place in the database and in the source document again. It would be very clear if any fact had been missed, because you would see an empty field in your template. It would be easy to check the data for errors because it's all there in one place looking much like the source document.

My ideal software would have an easy way to identify individuals in the document as individuals who are already in the database, or as new people to add. You wouldn't have to come up with some elaborate identification scheme. If you later decided that the source didn't refer to that person, you should just be able to unlink that identification without having to change anything else.

The software should make some sensible assumptions about how the information in the source document fits together and build the lineage links for you on that basis - but you should be able to review and override those assumptions if you wish. It should also be easy to add in any information from the source that is not standard for the template. You should also be able to add information from other sources that don't come neatly presented in a form.

It seems like a big ask, which makes this post seem like a rant... but guess what? Just under two weeks ago I stumbled across a genealogy package I hadn't heard of before. It promised source-based data entry along the lines I describe. I've been having a ball playing with the trial version for nearly two weeks now. While it's not perfect, I think it's interesting enough to write about in my next post...

That's one element of my ideal for genealogy software. Is there a genealogy software feature that seems so obvious and sensible to you that you just can't understand why anyone hasn't done it (to your satisfaction) before?!