<?xml version='1.0' encoding='utf-8'?>
<?xml-stylesheet type="text/xsl" href="/v2/static/oai2.xsl"?>
<OAI-PMH xmlns="http://www.openarchives.org/OAI/2.0/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/ http://www.openarchives.org/OAI/2.0/OAI-PMH.xsd">
  <responseDate>2026-10-09T21:04:30Z</responseDate>
  <request identifier="oai:figshare.com:article/33941767" metadataPrefix="oai_dc" verb="GetRecord">https://api.figshare.com/v2/oai</request>
  <GetRecord>
    <record>
      <header>
        <identifier>oai:figshare.com:article/33941767</identifier>
        <datestamp>2026-09-19T08:29:42Z</datestamp>
        <setSpec>category_24193</setSpec>
        <setSpec>category_24202</setSpec>
        <setSpec>category_24184</setSpec>
        <setSpec>category_24382</setSpec>
        <setSpec>item_type_3</setSpec>
        <setSpec>month_year_09_2026</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance"  xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xmlns:dc="http://purl.org/dc/elements/1.1/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:title>DEAR-OWL validation scripts and data</dc:title>
          <dc:creator>Daisuke Tsugama (9977096)</dc:creator>
          <dc:creator>Kota Kambara (17862556)</dc:creator>
          <dc:subject>Genomics and transcriptomics</dc:subject>
          <dc:subject>Statistical and quantitative genetics</dc:subject>
          <dc:subject>Bioinformatic methods development</dc:subject>
          <dc:subject>Plant cell and molecular biology</dc:subject>
          <dc:subject>web browser application</dc:subject>
          <dc:subject>RNA-seq</dc:subject>
          <dc:subject>transcriptome</dc:subject>
          <dc:subject>differential expression analysis</dc:subject>
          <dc:subject>differentially expressed genes</dc:subject>
          <dc:subject>DESeq2</dc:subject>
          <dc:subject>EdgeR</dc:subject>
          <dc:subject>Grass Expression Atlas</dc:subject>
          <dc:subject>Motif analysis</dc:subject>
          <dc:subject>HOMER</dc:subject>
          <dc:description>&lt;p&gt;&lt;strong&gt;Data Organization and File Naming Convention&lt;/strong&gt;&lt;/p&gt;&lt;p&gt;To ensure maximum computational reproducibility, all validation scripts, input records, and generated output files are contained within a single flat directory structure. This design allows the workflows to be executed immediately without requiring complex relative directory path configurations. Users can easily identify and utilize the datasets based on the following standardized file naming conventions:&lt;/p&gt;&lt;p&gt;&lt;br&gt;&lt;/p&gt;&lt;ul&gt;    &lt;li&gt;&lt;strong&gt;Execution &amp; Processing Scripts:&lt;/strong&gt;        &lt;ul&gt;            &lt;li&gt;&lt;code&gt;validation_script.txt&lt;/code&gt;: The primary shell one-liner script used to automate the sequential processing, file renaming, and tool executions.&lt;/li&gt;            &lt;li&gt;&lt;code&gt;*.py&lt;/code&gt;: Python scripts responsible for statistical analysis and generating performance/consensus plots (called internally by the shell script).&lt;/li&gt;        &lt;/ul&gt;    &lt;/li&gt;    &lt;li&gt;&lt;strong&gt;DEAR-OWL Output Lists:&lt;/strong&gt;        &lt;ul&gt;            &lt;li&gt;&lt;code&gt;uploaded*.csv&lt;/code&gt;: The DEG consensus lists containing fold-change values and p-values exported from DEAR-OWL. Files derived from the lightweight engine are explicitly distinguished by the suffix &lt;code&gt;_edgeR-like.csv&lt;/code&gt;.&lt;/li&gt;        &lt;/ul&gt;    &lt;/li&gt;    &lt;li&gt;&lt;strong&gt;Sequence Data (FASTA):&lt;/strong&gt;        &lt;ul&gt;            &lt;li&gt;&lt;code&gt;uploaded*.fa&lt;/code&gt;: The 500-bp upstream promoter sequences corresponding to the target DEG lists.&lt;/li&gt;        &lt;/ul&gt;    &lt;/li&gt;    &lt;li&gt;&lt;strong&gt;Reference DEG Datasets:&lt;/strong&gt;        &lt;ul&gt;            &lt;li&gt;&lt;code&gt;*brushed*.txt&lt;/code&gt;: Reference DEG lists previously identified and published in Yoon et al. (2024, &lt;em&gt;Plant Physiol Biochem&lt;/em&gt;).&lt;/li&gt;            &lt;li&gt;&lt;code&gt;PRJNA*.txt&lt;/code&gt;: Reference DEG lists corresponding to the &lt;em&gt;Cenchrus americanus&lt;/em&gt; samples from Qazi et al. (2025, &lt;em&gt;Biosci Biotechnol Biochem&lt;/em&gt;), also archived in Tsugama et al. (2024, Figshare).&lt;/li&gt;        &lt;/ul&gt;    &lt;/li&gt;    &lt;li&gt;&lt;strong&gt;HOMER Analysis Outputs:&lt;/strong&gt;        &lt;ul&gt;            &lt;li&gt;&lt;code&gt;Homer*&lt;/code&gt; (Directories/Folders): Output directories containing the definitive motif discovery tables, enrichment scores, and sequence logos generated by the HOMER suite.&lt;/li&gt;        &lt;/ul&gt;    &lt;/li&gt;    &lt;li&gt;&lt;strong&gt;Intermediate and Outcome Files:&lt;/strong&gt;        &lt;ul&gt;            &lt;li&gt;Other &lt;code&gt;*.png&lt;/code&gt;, &lt;code&gt;*.txt&lt;/code&gt;, and miscellaneous formats represent the raw input matrices, mid-stage transformation records, and definitive output results.&lt;/li&gt;        &lt;/ul&gt;    &lt;/li&gt;&lt;/ul&gt;&lt;p&gt;&lt;br&gt;&lt;/p&gt;&lt;p&gt;The specific genetic background, treatment conditions, and sample identifiers (e.g., QM2, VIP1-GFP) are explicitly included in the prefix of each respective filename for clear cross-referencing.&lt;/p&gt;</dc:description>
          <dc:date>2026-09-19T08:29:42Z</dc:date>
          <dc:type>Dataset</dc:type>
          <dc:type>Dataset</dc:type>
          <dc:identifier>10.6084/m9.figshare.33941767.v2</dc:identifier>
          <dc:relation>https://figshare.com/articles/dataset/DEAR-OWL_validation_scripts_and_data/33941767</dc:relation>
          <dc:rights>CC BY 4.0</dc:rights>
        </oai_dc:dc>
      </metadata>
    </record>
  </GetRecord>
</OAI-PMH>
