Gist ccffb77a3cba5716d7a983e6c009de95
✓ Published0🌍 Public
NNpfn
Last edited Feb 8, 2019
Created on Feb 8, 2019
This example shows a Perl script that extracts protein IDs and functional information from CEGMA output files by matching KOG identifiers. The code reads two input files, parses each line with regular expressions, and writes matching KOG entries with their descriptions to a new output file. The implementation relies on standard Perl file handling and pattern matching with the `/KOG\d{4}/` and `/^\[.+\]\s(KOG\d{4})\s(.*)$/` regex patterns, without using any external visualization libraries or APIs.
AI-generated description