<div>I think we are in agreement here, but my point is adding a read cache to Lustre (which is feature I rely heavily on) isn&#39;t the fundmental problem, it just exposed it. The fundamental problem is with ext[34] performance on larger disks.</div>

<div></div>
<div>Jeremy<br><br></div>
<div class="gmail_quote">On Wed, Apr 20, 2011 at 8:57 AM, Kevin Van Maren <span dir="ltr">&lt;<a href="mailto:kevin.van.maren@oracle.com">kevin.van.maren@oracle.com</a>&gt;</span> wrote:<br>
<blockquote style="BORDER-LEFT: #ccc 1px solid; MARGIN: 0px 0px 0px 0.8ex; PADDING-LEFT: 1ex" class="gmail_quote">Yes, difficulty finding free disk space can also be a problem, but I<br>could not recall big changes in<br>
how that worked since 1.6, other than memory pressure from the read<br>cache pushing out the bitmaps.<br>See <a href="http://jira.whamcloud.com/browse/LU-15" target="_blank">http://jira.whamcloud.com/browse/LU-15</a><br><font color="#888888"><br>
Kevin<br></font>
<div>
<div></div>
<div class="h5"><br><br>James Rose wrote:<br>&gt; Hi Kevin,<br>&gt;<br>&gt; Thanks for the suggestion. I will try this out.<br>&gt;<br>&gt; For the moment it seems that it may be disk space related. I have<br>&gt; removed some data from the file system. Performance returned to where I<br>
&gt; would expect it to be as space freed up (currently at 83% full). Since<br>&gt; free space I have seen two messages on an OSS where the number of<br>&gt; threads is tuned to the amount of RAM in the host and six on an OSS that<br>
&gt; has the number of threads set higher than it should. This is a much<br>&gt; better situation than the steady stream I was experiencing last night.<br>&gt; Maybe disabling the read cache will remove the last few.<br>
&gt;<br>&gt; I am still very curious what the rapid small reads seen when writing are<br>&gt; as this showed up while mounted ldiskfs so not doing regular lustre<br>&gt; operations at all.<br>&gt;<br>&gt; Thanks again for your help,<br>
&gt;<br>&gt; James.<br>&gt;<br>&gt;<br>&gt; On Wed, 2011-04-20 at 08:48 -0300, Kevin Van Maren wrote:<br>&gt;<br>&gt;&gt; First guess is the increased memory pressure caused by the Lustre 1.8<br>&gt;&gt; read cache. Many times &quot;slow&quot; messages are caused by memory<br>
&gt;&gt; allocatons taking a long time.<br>&gt;&gt;<br>&gt;&gt; You could try disabling the read cache and see if that clears up the<br>&gt;&gt; slow messages.<br>&gt;&gt;<br>&gt;&gt; Kevin<br>&gt;&gt;<br>&gt;&gt;<br>&gt;&gt; On Apr 20, 2011, at 4:29 AM, James Rose &lt;<a href="mailto:James.Rose@framestore.com">James.Rose@framestore.com</a>&gt;<br>
&gt;&gt; wrote:<br>&gt;&gt;<br>&gt;&gt;<br>&gt;&gt;&gt; Hi<br>&gt;&gt;&gt;<br>&gt;&gt;&gt; We have been experiencing degraded performance for a few days on a<br>&gt;&gt;&gt; fresh install of lustre 1.8.5 (on RHEL5 using sun ext4 rpms). The<br>
&gt;&gt;&gt; initial bulk load of the data will be fine but once in use for a<br>&gt;&gt;&gt; while writes become very slow to individual ost. This will block io<br>&gt;&gt;&gt; for a few minutes and then carry on as normal. The slow writes will<br>
&gt;&gt;&gt; then move to another ost. This can be seen in iostat and many slow<br>&gt;&gt;&gt; IO messages will be seen in the logs (example included)<br>&gt;&gt;&gt;<br>&gt;&gt;&gt; The osts are between 87 90 % full. Not ideal but has not caused any<br>
&gt;&gt;&gt; issues running 1.6.7.2 on the same hardware.<br>&gt;&gt;&gt;<br>&gt;&gt;&gt; The osts are RAID6 on external raid chassis (Infortrend). Each ost<br>&gt;&gt;&gt; is 5.4T (small). The server is Dual AMD (4 cores). 16G Ram. Qlogic<br>
&gt;&gt;&gt; FC HBA.<br>&gt;&gt;&gt;<br>&gt;&gt;&gt; I mounted the osts as ldiskfs and tried a few write tests. These<br>&gt;&gt;&gt; also show the same behaviour.<br>&gt;&gt;&gt;<br>&gt;&gt;&gt; While the write operation is blocked there will be hundreds of read<br>
&gt;&gt;&gt; tps and a very small kb/s read from the raid but now writes. As<br>&gt;&gt;&gt; soon as this completes writes will go through at a more expected<br>&gt;&gt;&gt; speed.<br>&gt;&gt;&gt;<br>&gt;&gt;&gt; Any idea what is going on?<br>
&gt;&gt;&gt;<br>&gt;&gt;&gt; Many thanks<br>&gt;&gt;&gt;<br>&gt;&gt;&gt; James.<br>&gt;&gt;&gt;<br>&gt;&gt;&gt; Example error messages:<br>&gt;&gt;&gt;<br>&gt;&gt;&gt; Apr 20 04:53:04 oss5r-mgmt kernel: LustreError: dumping log to /tmp/<br>
&gt;&gt;&gt; lustre-log.1303271584.3935<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: rho-OST0012: slow quota<br>&gt;&gt;&gt; init 286s due to heavy IO load<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: rho-OST0012: slow journal<br>
&gt;&gt;&gt; start 39s due to heavy IO load<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: Skipped 39 previous<br>&gt;&gt;&gt; similar messages<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: rho-OST0012: slow<br>
&gt;&gt;&gt; brw_start 39s due to heavy IO load<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: Skipped 38 previous<br>&gt;&gt;&gt; similar messages<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: rho-OST0012: slow journal<br>
&gt;&gt;&gt; start 133s due to heavy IO load<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: Skipped 44 previous<br>&gt;&gt;&gt; similar messages<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: rho-OST0012: slow<br>
&gt;&gt;&gt; brw_start 133s due to heavy IO load<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: Skipped 44 previous<br>&gt;&gt;&gt; similar messages<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: rho-OST0012: slow journal<br>
&gt;&gt;&gt; start 236s due to heavy IO load<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: rho-OST0012: slow i_mutex<br>&gt;&gt;&gt; 40s due to heavy IO load<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: Skipped 2 previous<br>
&gt;&gt;&gt; similar messages<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: Skipped 6 previous<br>&gt;&gt;&gt; similar messages<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: rho-OST0012: slow i_mutex<br>
&gt;&gt;&gt; 277s due to heavy IO load<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: rho-OST0012: slow<br>&gt;&gt;&gt; direct_io 286s due to heavy IO load<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: Skipped 3 previous<br>
&gt;&gt;&gt; similar messages<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: rho-OST0012: slow journal<br>&gt;&gt;&gt; start 285s due to heavy IO load<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: Skipped 1 previous<br>
&gt;&gt;&gt; similar message<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: rho-OST0012: slow<br>&gt;&gt;&gt; commitrw commit 285s due to heavy IO load<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: Skipped 1 previous<br>
&gt;&gt;&gt; similar message<br>&gt;&gt;&gt; Apr 20 04:53:40 oss5r-mgmt kernel: Lustre: rho-OST0012: slow parent<br>&gt;&gt;&gt; lock 236s due to heavy IO load<br>&gt;&gt;&gt;<br>&gt;&gt;&gt;<br>&gt;&gt;&gt; _______________________________________________<br>
&gt;&gt;&gt; Lustre-discuss mailing list<br>&gt;&gt;&gt; <a href="mailto:Lustre-discuss@lists.lustre.org">Lustre-discuss@lists.lustre.org</a><br>&gt;&gt;&gt; <a href="http://lists.lustre.org/mailman/listinfo/lustre-discuss" target="_blank">http://lists.lustre.org/mailman/listinfo/lustre-discuss</a><br>
&gt;&gt;&gt;<br>&gt;&gt; _______________________________________________<br>&gt;&gt; Lustre-discuss mailing list<br>&gt;&gt; <a href="mailto:Lustre-discuss@lists.lustre.org">Lustre-discuss@lists.lustre.org</a><br>&gt;&gt; <a href="http://lists.lustre.org/mailman/listinfo/lustre-discuss" target="_blank">http://lists.lustre.org/mailman/listinfo/lustre-discuss</a><br>
&gt;&gt;<br>&gt;<br>&gt;<br>&gt;<br><br>_______________________________________________<br>Lustre-discuss mailing list<br><a href="mailto:Lustre-discuss@lists.lustre.org">Lustre-discuss@lists.lustre.org</a><br><a href="http://lists.lustre.org/mailman/listinfo/lustre-discuss" target="_blank">http://lists.lustre.org/mailman/listinfo/lustre-discuss</a><br>
</div></div></blockquote></div><br>