提交 · e13d224a171ca31556118081225ebfc4b6018142 · OpenXiangShan / XiangShan

14 12月, 2021 5 次提交

Y

difftest: move sc_valid to AtomicsUnit (#1350) · e13d224a
由 Yinan Xu 提交于 12月 14, 2021

e13d224a

dp2: out.bits does not depend on lsq.canAccept (#1352) · 74ca315b

由 Yinan Xu 提交于 12月 14, 2021

This commit optimizes Dispatch2Rs timing by ignoring lsq.canAccept
when sending bits to reservation stations.

74ca315b

Optimize IFU and PreDecode timing (#1347) · 2a3050c2

由 Jay 提交于 12月 14, 2021

* ICache: add ReplacePipe for Probe & Release

* remove ProbeUnit

* Probe & Release enter ReplacePipe

* fix bugs when running Linux on MinimalConfig

* TODO: set conflict for ReplacePipe

* ICache: fix ReplacePipe invalid write bug

* chores: code clean up

* IFU: optimize timing

* PreDecode: separate into 2 module for timing optimization

* IBuffer: add enqEnable to replace valid for timing

* IFU/ITLB: optimize timing

* IFU: calculate cut_ptr in f1

* TLB: send req in f1 and wait resp in f2

* ICacheMainPipe: add tlb miss logic in s0

* Optimize IFU timing

* IFU: fix lastHalfRVI bug

* IFU: fix performance bug

* IFU: optimize MMIO commit timing

* IFU: optmize trigger timing and add frontendTrigger

* fix compile error

* IFU: fix mmio stuck bug

2a3050c2

Z

dcache: fix bug in ecc check (#1349) · dd95524e
由 zhanglinjuan 提交于 12月 14, 2021

dd95524e

csr: update mtval/stval according to the trap mode (#1344) · 7c071650

由 Yinan Xu 提交于 12月 14, 2021

This commit changes the condition to update mtval and stval.

According to the RISC-V spec, when a trap is taken into M/S-mode,
mtval/stval is either set to zero or written wrih exception-specific
information to assist software in handling the trap.

Previously in XiangShan, mtval/stval is updated depending on the
current priviledge mode, which is incorrect.

7c071650

13 12月, 2021 3 次提交

Optimize dcache timing (#1332) · 69790076

由 zhanglinjuan 提交于 12月 13, 2021

* MissQueue: loose merging condition to ease timing stress

* MissQueue: remove grant_beats

* MissQueue: compare block addr, not the whole addr bits

* dcache: optimize timing for generating ready to sbuffer
Co-authored-by: NWilliam Wang <zeweiwang@outlook.com>

69790076

Y
Merge pull request #1345 from OpenXiangShan/fix-soft-prefetch · 979fa9bc
由 Yinan Xu 提交于 12月 13, 2021
```
mem: fix soft prefetch
```
979fa9bc

SoC: insert more buffers into mmio path (#1329) · be340b14

由 Jiawei Lin 提交于 12月 13, 2021

* SoC: add axi4spliter

* pmp: add apply method to reduce loc

* pma: add PMA used in axi4's spliter

* Fix package import

* pma: re-write tl-pma, put tl-pma into AXI4Spliter

* pma: add memory mapped pma

* soc: rm dma port, rm axi4spliter, mv mmpma out of spliter

* csr: clear mstatus.mprv when mstatus.mpp != ModeM at xret

* csr: fix write mask for mstatus, mepc and sepc

This commit fixes the write mask for mstatus, mepc and sepc.

According to the RISC-V instruction manual, for RV64 systems,
the SXL and UXL fields are WARL fields that control the value of
XLEN for S-mode and U-mode, respectively. For RV64 systems, if
S-mode is not supported, then SXL is hardwired to zero. For RV64
systems, if U-mode is not supported, then UXL is hardwired to zero.

Besides, mepc[0] and sepc[0] should be hardwired to zero.

* wb,load: delay load fp for one cycle

* csr: add mconfigptr, but hardwire to 0 now

* bump huancun

* csr: add *BE to mstatusStruct which are hardwired to 0

* Remove unused files

* csr: fix bug of xret clear mprv

* bump difftest

* ci: add unit test, xret clear mstatus.mprv when xpp is not M

* bump ready-to-run

* mem,atomics: delay exception info for one cycle

* SoC: insert more buffers into mmio path

* SoC: insert buffer between l3_xbar and l3_banked_xbar

* Optimze l3->ddr path

* Bump huancun
Co-authored-by: NZhangZifei <zhangzifei20z@ict.ac.cn>
Co-authored-by: NYinan Xu <xuyinan@ict.ac.cn>
Co-authored-by: Nwangkaifan <wangkaifan@ict.ac.cn>

be340b14

12 12月, 2021 5 次提交

W

mem: replay soft prefetch if tlb miss · c707f0c8
由 William Wang 提交于 12月 12, 2021

c707f0c8

L2/L3: fix prefetch train address (#1339) · 459ad1b2

由 Jiawei Lin 提交于 12月 12, 2021

* L2/L3: fix prefetch train address

* HuanCun: update SRAMTemplate

* Config: Keep the client dir capacity of L3 twice the L2

* Bump huancun

459ad1b2

W

csr: add soft_prefetch_enable to smblockctl · d10a581e
由 William Wang 提交于 12月 12, 2021

d10a581e
W
mem: soft prefetch will not be replayed · 690158b0
由 William Wang 提交于 12月 12, 2021
```
Soft prefetch will be always marked as "load hit"
```
690158b0

csr: add vectored trap mode (#1343) · 68b89fcb

由 Yinan Xu 提交于 12月 12, 2021

All bits for stvec and mtvec are writable in XiangShan.

According to the RISC-V spec, {m,s}tvec[1:0] are MODE bits. When
MODE=Vectored, all synchronous exceptions into M/S mode cause the pc
to be set to the address in the BASE field, whereas interrupts cause
the pc to be set to the address in the BASE field plus four times
the interrupt cause number.

If XiangShan decides to not support vectored mode, {m,s}tvec[1:0]
should be hardwired to zero.

68b89fcb

11 12月, 2021 4 次提交

jump: set the LSB of the target to zero (#1342) · 1a389dfd

由 Yinan Xu 提交于 12月 11, 2021

According to RISC-V spec, for the JALR instruction, its target address
is obtained by adding the sign-extended 12-bit I-immediate to the
register rs1, then setting the least-significant bit of the result
to zero.

1a389dfd

Y

csr: delay fflags and dirty_fs for better timing (#1341) · 7181c0c1
由 Yinan Xu 提交于 12月 11, 2021

7181c0c1

mmu: timing optimization of ptwfilter's recv and issue & storeunit's mmio (#1326) · 2c2c1588

由 Lemover 提交于 12月 11, 2021

* TLB: when miss, regnext the req sent to ptw

* PTWFilter: timing optimzation of do_iss that ignore ptwResp's filter

* StoreUnit: logic optimization of from s2_mmio to s2_out_valid

* ptwfilter: when issue but filtered, clear the v bit

special case that
ptw.resp clear all the duplicate req when arrive to filter
ptw_resp is the RegNext of ptw.resp and it filters ptw.req
when ptw_resp filter the req but ptw.resp not filter the tlb_req to
stop do_enq, then the v bit of the req will not be cleared ever.

It will be more correct to fliter the entries and tlb_req with ptw_resp,
but the timing restriction says no. So just use the confusing trick
to slove the complicate corner case.

2c2c1588

core: delay csrCtrl for two cycles (#1336) · 6f688dac

由 Yinan Xu 提交于 12月 11, 2021

This commit adds DelayN(2) to some CSR-related signals, including
control bits to ITLB, DTLB, PTW, etc.

To avoid accessing the ITLB before control bits change, we also need
to delay the flush for two cycles. We assume branch misprediction or
memory violation does not cause csrCtrl to change.

6f688dac

10 12月, 2021 4 次提交

icache: support data/tag r/w op (#1337) · 70899835

由 William Wang 提交于 12月 10, 2021

* mem,cacheop: fix read data writeback

* mem,cacheop: rename cacheop state bits

These bits are different from w_*, s_* bits in cache

* mem: enable icache op feedback

* icache: update cache op implementation

* chore: remove cache op logic from XSCore.scala

70899835

W

dcache: fix lrsc_locked_block check (#1334) · 8b538b51
由 William Wang 提交于 12月 10, 2021

8b538b51

core: refactor hardware performance counters (#1335) · 1ca0e4f3

由 Yinan Xu 提交于 12月 10, 2021

This commit optimizes the coding style and timing for hardware
performance counters.

By default, performance counters are RegNext(RegNext(_)).

1ca0e4f3

bump huancun (#1322) · 1dc3a3a0

由 wakafa 提交于 12月 10, 2021

* bump huancun

* bump huancun

* Bump huancun
Co-authored-by: NLinJiawei <linjiawei20s@ict.ac.cn>

1dc3a3a0

09 12月, 2021 2 次提交

J

ICache: send ProbeAck when Probe NToN (#1331) · 1d4a76ae
由 Jay 提交于 12月 09, 2021

1d4a76ae

core: refactor writeback parameters (#1327) · 6ab6918f

由 Yinan Xu 提交于 12月 09, 2021

This commit adds WritebackSink and WritebackSource parameters for
multiple modules. These traits hide implementation details from
other modules by defining IO-related functions in modules.

By using WritebackSink, ROB is able to choose the writeback sources.
Now fflags and exceptions are connected from exe units to reduce write
ports and optimize timing.

Further optimizations on write-back to RS and better coding style to
be added later.

6ab6918f

08 12月, 2021 4 次提交

csr: add write mask to satp.ppn & xstatus.xs (#1323) · 705cbec3

由 Lemover 提交于 12月 08, 2021

* csr.satp: add r/w mask of ppn part

* ci: add unit test, satp should concern PADDRBITS

* csr.xstatus: XS field is ready-only

* bump ready-to-run

* bump ready-to-run, update nemu so

* fix typo

705cbec3

dcache: optimize refill block timing (#1320) · b36dd5fd

由 William Wang 提交于 12月 08, 2021

Now we RegNext(refill_req) for 1 cycle. It will provide more
time for refillShouldBeBlocked calcuation

b36dd5fd

Fix dcache probe (#1324) · 53e88463

由 William Wang 提交于 12月 08, 2021

* dcache: give probe the highest priority

* dcache: fix block probe logic

* dcache: give replace_req higher priority

53e88463

R

update f2_mmio update logic (#1325) · c0b2b8e9
由 rvcoresjw 提交于 12月 08, 2021

c0b2b8e9

07 12月, 2021 3 次提交
- W
  dcache: fix read data cache op (#1319) · b6358f8f
  由 William Wang 提交于 12月 07, 2021
```
* mem,cacheop: fix read data writeback

* mem,cacheop: rename cacheop state bits

These bits are different from w_*, s_* bits in cache
```
  b6358f8f
- J
  
  ICache: fix same vidx req rewrite bug (#1316) · 6cc2baa1
  由 Jay 提交于 12月 07, 2021
  
  6cc2baa1
- J
  
  DTS: add interrupt-controller into cpu (#1298) · 7ba24bbc
  由 Jiawei Lin 提交于 12月 07, 2021
  
  7ba24bbc
06 12月, 2021 9 次提交

J

ICache: fix probe pipe_req.ready bug (#1318) · c90cd2d1
由 Jay 提交于 12月 06, 2021

c90cd2d1
J

ICache: Release always send ReleaseAckData (#1317) · f8e8fe29
由 Jay 提交于 12月 06, 2021

f8e8fe29
L
Fix SRT16div bug with 0 remainder (#1315) · 2acd2853
由 Li Qianruo 提交于 12月 06, 2021
```
This bug occurs when rem is 0 and dividend is negative
Caused by a buggy rightshifter
```
2acd2853

Add pma checker for I/O device (#1300) · 98c71602

由 Jiawei Lin 提交于 12月 06, 2021

* SoC: add axi4spliter

* pmp: add apply method to reduce loc

* pma: add PMA used in axi4's spliter

* Fix package import

* pma: re-write tl-pma, put tl-pma into AXI4Spliter

* pma: add memory mapped pma

* soc: rm dma port, rm axi4spliter, mv mmpma out of spliter

* Remove unused files

* update dma pma check port at SimTop.scala; update pll lock defalt value to 1
Co-authored-by: NZhangZifei <zhangzifei20z@ict.ac.cn>
Co-authored-by: Nrvcoresjw <shangjiawei@rvcore.com>

98c71602

W

mdp: fix valid_sram write assertion (#1306) · 0fbe42c4
由 William Wang 提交于 12月 06, 2021

0fbe42c4
J

ICache: fix set conflict condition (#1313) · 92acb6b9
由 Jay 提交于 12月 06, 2021

92acb6b9

Updated to priv 1.12 (#1301) · 7d9edc86

由 Lemover 提交于 12月 06, 2021

* csr: clear mstatus.mprv when mstatus.mpp != ModeM at xret

* csr: add mconfigptr, but hardwire to 0 now

* csr: add *BE to mstatusStruct which are hardwired to 0

* csr: fix bug of xret clear mprv

* ci: add unit test, xret clear mstatus.mprv when xpp is not M

* bump ready-to-run

7d9edc86

arbiter: better balance among function units (#1305) · d415b7f7

由 Yinan Xu 提交于 12月 06, 2021

This commit changes the splitN algorithm for the write-back arbiter.

Previously we split the function units as follows:
(FU0 FU1 FU2) (FU3 FU4 FU5).
However, this strategy tends to group the function units with the same
type into the same arbiter and may cause performance loss.

In this commit, we change the strategy to: (FU0 FU2 FU4) (FU1 FU3 FU5).

d415b7f7

rs: optimize issue grant timing with age (#1312) · 2234af84

由 Yinan Xu 提交于 12月 06, 2021

This commit optimizes the issue grant timing when age is enabled.
Select from age and SelectPolicy are processed parallely.

2234af84

05 12月, 2021 1 次提交
- W
  bump huancun (#1299) · d600fd3a
  由 wakafa 提交于 12月 05, 2021
```
* bump huancun
```
  d600fd3a

OpenXiangShan / XiangShan 10 个月 前同步成功

OpenXiangShan / XiangShan
10 个月前同步成功