What I will do in this space is give you the types of data that are at the high level hierarchy of my human genome database. These data types will likely be applicable to most eukaryotic gene annotation efforts. Prokaryote gene annotation is not something with which I have much experience and so should not offer suggestions, other than to say I think the environment in which the organism normally lives and was isolated is important. I have a lot of gene expression and protein expression (proteomics) data that I use to ascertain function or candidacy for further experiments in our lab, but I rely on all types of data to make that call.
As I work with human data, I am not so concerned with mapping sequence by BLAST. I simply note the genome build and the associated gene coordinates, ignoring exon coords (for now).
[?]
Lastly, I have a free-text "knowledge" field where I enter info from lab meeting, abstracts, etc. (typically with reference).
In addition, I have a metabolite database where I link small molecule to gene.
That list should give you some good ideas on what to collect in order to more confidently describe the potential function of a gene and its encoded protein.
If these are something like microarray probes, you might consider answers to this question:
A: Pipeline To Map 60-Mers To Genes