Sequence Analysis in a Nutshell - A Guide to Common Tools & Databases

Author:   Scott Markel ,  Darryl Leon
Publisher:   O'Reilly Media
ISBN:  

9780596004941


Pages:   304
Publication Date:   04 March 2003
Format:   Paperback
Availability:   Out of print, replaced by POD   Availability explained
We will order this item for you from a manufatured on demand supplier.

Our Price $79.07 Quantity:  
Add to Cart

Share |

Sequence Analysis in a Nutshell - A Guide to Common Tools & Databases


Add your own review!

Overview

Gene sequence data is the most abundant type of data available, and if you're interested in analyzing it, you'll find a wealth of computational methods and tools to help you. In fact, finding the data is not the challenge at all; rather it is dealing with the plethora of flat file formats used to process the sequence entries and trying to remember what their specific field codes mean. This book is a handy resource, as well as an invaluable reference, for anyone who needs to know about the practical aspects and mechanics of sequence analysis. This reference pulls together all of the vital information about the most commonly used databases, analytical tools, and tables used in sequence analysis. The book is partitioned into three fundamental areas to help you maximize your use of the content. The first section, ""Databases"" contains examples of flatfiles from key databases (GenBank, EMBL, SWISS-PROT), the definitions of the codes or fields used in each database, and the sequence feature types/terms and qualifiers for the nucleotide and protein databases. The second section, """"Tools"""" provides the command line syntax for popular applications such as ReadSeq, MEME/MAST, BLAST, ClustalW, and the EMBOSS suite of analytical tools. The third section, """"Appendixes"""" concentrates on information essential to understanding the individual components that make up a biological sequence. The tables in this section include nucleotide and protein codes, genetic codes, as well as other relevant information. Written in O'Reilly's straightforward """"Nutshell"""" format, this book draws together essential information for bioinformaticians in industry and academia, as well as for students. ""

Full Product Details

Author:   Scott Markel ,  Darryl Leon
Publisher:   O'Reilly Media
Imprint:   O'Reilly Media
Dimensions:   Width: 16.60cm , Height: 1.80cm , Length: 22.90cm
Weight:   0.410kg
ISBN:  

9780596004941


ISBN 10:   059600494
Pages:   304
Publication Date:   04 March 2003
Audience:   College/higher education ,  Professional and scholarly ,  Undergraduate ,  Postgraduate, Research & Scholarly
Format:   Paperback
Publisher's Status:   Active
Availability:   Out of print, replaced by POD   Availability explained
We will order this item for you from a manufatured on demand supplier.

Table of Contents

Preface I. Data Formats 1. FASTA Format NCBI's Sequence Identifier Syntax NCBI's Non-Redundant Database Syntax References 2. GenBank/EMBL/DDBJ Example Flat Files GenBank Example Flat File DDBJ Example Flat File GenBank/DDBJ Field Definitions EMBL Example Flat File EMBL Field Definitions DDBJ/EMBL/GenBank Feature Table References 3. SWISS-PROT SWISS-PROT Example Flat File SWISS-PROT Field Definitions SWISS-PROT Feature Table References 4. Pfam Pfam Example Flat File Pfam Field Definitions References 5. PROSITE PROSITE Example Flat File PROSITE Field Definitions References II. Tools 6. Readseq Supported Formats Command-Line Options References 7. BLAST formatdb blastall megablast blastpgp PSI-BLAST PHI-BLAST bl2seq References 8. BLAT Command-Line Options References 9. ClustalW Command-Line Options References 10. HMMER hmmalign hmmbuild hmmcalibrate hmmconvert hmmemit hmmfetch hmmindex hmmpfam hmmsearch References 11. MEME/MAST MEME MAST References 12. EMBOSS Common Themes List of All EMBOSS Programs Details of EMBOSS Programs References III. Appendixes A. Nucleotide and Amino Acid Tables B. Genetic Codes C. Resources D. Future Plans Index

Reviews

Author Information

Scott Markel is a Principal Software Architect at LION bioscience Inc., where he is responsible for providing architectural direction in the development of software for the life sciences, including the use and development of standards. He is a co-chair of the Life Sciences Research Domain Task Force of the Object Management Group, and also chairs the LSR's Architecture and Roadmap Working Group. Prior to working at LION, Scott worked at NetGenics, Johnson & Johnson Pharmaceutical Research & Development, and Sarnoff Corporation. He has a Ph.D. in mathematics from the University of Wisconsin-Madison. When Scott's not working or writing he enjoys spending time with his wife and kids, reading European history books, and just enjoying life in sunny San Diego. Darryl Leon is a Principal Scientific Architect at LION bioscience Inc., where he is responsible for providing scientific direction in the development of software for the life sciences. Prior to working at LION, Darryl worked at NetGenics, DoubleTwist, and Genset. He has taught at California Polytechnic State University, San Luis Obispo, and currently teaches a bioinformatics class at U.C. Santa Cruz Extension and U.C. San Diego Extension. He is also a member of the Bioinformatics Advisory Committee at U.C. San Diego Extension. Darryl has a Ph. D. in biochemistry from the University of California, San Diego and did his postdoctoral research at the University of California, Santa Cruz.

Tab Content 6

Author Website:  

Customer Reviews

Recent Reviews

No review item found!

Add your own review!

Countries Available

All regions
Latest Reading Guide

Aorrng

Shopping Cart
Your cart is empty
Shopping cart
Mailing List