Nature Genetics 2004-01-01

Complete sequencing and characterization of 21,243 full-length human cDNAs.

Toshio Ota, Yutaka Suzuki, Tetsuo Nishikawa, Tetsuji Otsuki, Tomoyasu Sugiyama, Ryotaro Irie, Ai Wakamatsu, Koji Hayashi, Hiroyuki Sato, Keiichi Nagai, Kouichi Kimura, Hiroshi Makita, Mitsuo Sekine, Masaya Obayashi, Tatsunari Nishi, Toshikazu Shibahara, Toshihiro Tanaka, Shizuko Ishii, Jun-ichi Yamamoto, Kaoru Saito, Yuri Kawai, Yuko Isono, Yoshitaka Nakamura, Kenji Nagahari, Katsuhiko Murakami, Tomohiro Yasuda, Takao Iwayanagi, Masako Wagatsuma, Akiko Shiratori, Hiroaki Sudo, Takehiko Hosoiri, Yoshiko Kaku, Hiroyo Kodaira, Hiroshi Kondo, Masanori Sugawara, Makiko Takahashi, Katsuhiro Kanda, Takahide Yokoi, Takako Furuya, Emiko Kikkawa, Yuhi Omura, Kumi Abe, Kumiko Kamihara, Naoko Katsuta, Kazuomi Sato, Machiko Tanikawa, Makoto Yamazaki, Ken Ninomiya, Tadashi Ishibashi, Hiromichi Yamashita, Katsuji Murakawa, Kiyoshi Fujimori, Hiroyuki Tanai, Manabu Kimata, Motoji Watanabe, Susumu Hiraoka, Yoshiyuki Chiba, Shinichi Ishida, Yukio Ono, Sumiyo Takiguchi, Susumu Watanabe, Makoto Yosida, Tomoko Hotuta, Junko Kusano, Keiichi Kanehori, Asako Takahashi-Fujii, Hiroto Hara, Tomo-o Tanase, Yoshiko Nomura, Sakae Togiya, Fukuyo Komai, Reiko Hara, Kazuha Takeuchi, Miho Arita, Nobuyuki Imose, Kaoru Musashino, Hisatsugu Yuuki, Atsushi Oshima, Naokazu Sasaki, Satoshi Aotsuka, Yoko Yoshikawa, Hiroshi Matsunawa, Tatsuo Ichihara, Namiko Shiohata, Sanae Sano, Shogo Moriya, Hiroko Momiyama, Noriko Satoh, Sachiko Takami, Yuko Terashima, Osamu Suzuki, Satoshi Nakagawa, Akihiro Senoh, Hiroshi Mizoguchi, Yoshihiro Goto, Fumio Shimizu, Hirokazu Wakebe, Haretsugu Hishigaki, Takeshi Watanabe, Akio Sugiyama, Makoto Takemoto, Bunsei Kawakami, Masaaki Yamazaki, Koji Watanabe, Ayako Kumagai, Shoko Itakura, Yasuhito Fukuzumi, Yoshifumi Fujimori, Megumi Komiyama, Hiroyuki Tashiro, Akira Tanigami, Tsutomu Fujiwara, Toshihide Ono, Katsue Yamada, Yuka Fujii, Kouichi Ozaki, Maasa Hirao, Yoshihiro Ohmori, Ayako Kawabata, Takeshi Hikiji, Naoko Kobatake, Hiromi Inagaki, Yasuko Ikema, Sachiko Okamoto, Rie Okitani, Takuma Kawakami, Saori Noguchi, Tomoko Itoh, Keiko Shigeta, Tadashi Senba, Kyoka Matsumura, Yoshie Nakajima, Takae Mizuno, Misato Morinaga, Masahide Sasaki, Takushi Togashi, Masaaki Oyama, Hiroko Hata, Manabu Watanabe, Takami Komatsu, Junko Mizushima-Sugano, Tadashi Satoh, Yuko Shirai, Yukiko Takahashi, Kiyomi Nakagawa, Koji Okumura, Takahiro Nagase, Nobuo Nomura, Hisashi Kikuchi, Yasuhiko Masuho, Riu Yamashita, Kenta Nakai, Tetsushi Yada, Yusuke Nakamura, Osamu Ohara, Takao Isogai, Sumio Sugano

文献索引:Nat. Genet. 36 , 40-5, (2004)

全文:HTML全文

摘要

As a base for human transcriptome and functional genomics, we created the "full-length long Japan" (FLJ) collection of sequenced human cDNAs. We determined the entire sequence of 21,243 selected clones and found that 14,490 cDNAs (10,897 clusters) were unique to the FLJ collection. About half of them (5,416) seemed to be protein-coding. Of those, 1,999 clusters had not been predicted by computational methods. The distribution of GC content of nonpredicted cDNAs had a peak at approximately 58% compared with a peak at approximately 42%for predicted cDNAs. Thus, there seems to be a slight bias against GC-rich transcripts in current gene prediction procedures. The rest of the cDNAs unique to the FLJ collection (5,481) contained no obvious open reading frames (ORFs) and thus are candidate noncoding RNAs. About one-fourth of them (1,378) showed a clear pattern of splicing. The distribution of GC content of noncoding cDNAs was narrow and had a peak at approximately 42%, relatively low compared with that of protein-coding cDNAs.


相关化合物

  • 3-磷酸甘油醛脱氢酶
  • 脱氧鸟苷酸激酶
  • 谷丙转氨酶 来源于...
  • 肌氨酸氧化酶
  • 醛缩酶
  • 脂肪酶
  • β-半乳糖苷酶
  • L-天门冬酰胺酶
  • D-核酮糖-5-磷酸-3-...
  • 肌激酶 来源于兔肌...

相关文献:

The complete sequence of a full length cDNA for human liver glyceraldehyde-3-phosphate dehydrogenase: evidence for multiple mRNA species.

1984-12-11

[Nucleic Acids Res. 12(23) , 9179-89, (1984)]

Quantitative phosphoproteomic analysis of T cell receptor signaling reveals system-wide modulation of protein-protein interactions.

2009-01-01

[Sci. Signal. 2 , ra46, (2009)]

The glyceraldehyde 3 phosphate dehydrogenase gene family: structure of a human cDNA and of an X chromosome linked pseudogene; amazing complexity of the gene family in mouse.

1984-11-01

[EMBO J. 3(11) , 2627-33, (1984)]

Initial characterization of the human central proteome.

2011-01-01

[BMC Syst. Biol. 5 , 17, (2011)]

The DNA sequence and biology of human chromosome 19.

2004-04-01

[Nature 428(6982) , 529-35, (2004)]

更多文献...