ftmemsim-valgrind

mirror of https://github.com/Zenithsiz/ftmemsim-valgrind.git synced 2026-02-10 13:40:25 +00:00

Author	SHA1	Message	Date
Mark Wielaard	461cc5c003	Cleanup GPL header address notices by using http://www.gnu.org/licenses/ Sync VEX/LICENSE.GPL with top-level COPYING file. We used 3 different addresses for writing to the FSF to receive a copy of the GPL. Replace all different variants with an URL <http://www.gnu.org/licenses/>. The following files might still have some slightly different (L)GPL copyright notice because they were derived from other programs: - files under coregrind/m_demangle which come from libiberty: cplus-dem.c, d-demangle.c, demangle.h, rust-demangle.c, safe-ctype.c and safe-ctype.h - coregrind/m_demangle/dyn-string.[hc] derived from GCC. - coregrind/m_demangle/ansidecl.h derived from glibc. - VEX files for FMA detived from glibc: host_generic_maddf.h and host_generic_maddf.c - files under coregrin/m_debuginfo derived from LZO: lzoconf.h, lzodefs.h, minilzo-inl.c and minilzo.h - files under coregrind/m_gdbserver detived from GDB: gdb/signals.h, inferiors.c, regcache.c, regcache.h, regdef.h, remote-utils.c, server.c, server.h, signals.c, target.c, target.h and utils.c Plus the following test files: - none/tests/ppc32/testVMX.c derived from testVMX. - ppc tests derived from QEMU: jm-insns.c, ppc64_helpers.h and test_isa_3_0.c - tests derived from bzip2 (with embedded GPL text in code): hackedbz2.c, origin5-bz2.c, varinfo6.c - tests detived from glibc: str_tester.c, pth_atfork1.c - test detived from GCC libgomp: tc17_sembar.c - performance tests derived from bzip2 or tinycc (with embedded GPL text in code): bz2.c, test_input_for_tinycc.c and tinycc.c	2019-05-26 20:07:51 +02:00
Julian Seward	50bb127b1d	Bug 402781 - Redo the cache used to process indirect branch targets. [This commit contains an implementation for all targets except amd64-solaris and x86-solaris, which will be completed shortly.] In the baseline simulator, jumps to guest code addresses that are not known at JIT time have to be looked up in a guest->host mapping table. That means: indirect branches, indirect calls and most commonly, returns. Since there are huge numbers of these (often 10+ million/second) the mapping mechanism needs to be extremely cheap. Currently, this is implemented using a direct-mapped cache, VG_(tt_fast), with 2^15 (guest_addr, host_addr) pairs. This is queried in handwritten assembly in VG_(disp_cp_xindir) in dispatch-<arch>-<os>.S. If there is a miss in the cache then we fall back out to C land, and do a slow lookup using VG_(search_transtab). Given that the size of the translation table(s) in recent years has expanded significantly in order to keep pace with increasing application sizes, two bad things have happened: (1) the cost of a miss in the fast cache has risen significantly, and (2) the miss rate on the fast cache has also increased significantly. This means that large (~ one-million-basic-blocks-JITted) applications that run for a long time end up spending a lot of time in VG_(search_transtab). The proposed fix is to increase associativity of the fast cache, from 1 (direct mapped) to 4. Simulations of various cache configurations using indirect-branch traces from a large application show that is the best of various configurations. In an extreme case with 5.7 billion indirect branches: * The increase of associativity from 1 way to 4 way, whilst keeping the overall cache size the same (32k guest/host pairs), reduces the miss rate by around a factor of 3, from 4.02% to 1.30%. * The use of a slightly better hash function than merely slicing off the bottom 15 bits of the address, reduces the miss rate further, from 1.30% to 0.53%. Overall the VG_(tt_fast) miss rate is almost unchanged on small workloads, but reduced by a factor of up to almost 8 on large workloads. By implementing each (4-entry) cache set using a move-to-front scheme in the case of hits in ways 1, 2 or 3, the vast majority of hits can be made to happen in way 0. Hence the cost of having this extra associativity is almost zero in the case of a hit. The improved hash function costs an extra 2 ALU shots (a shift and an xor) but overall this seems performance neutral to a win.	2019-01-25 09:14:56 +01:00
Ivo Raisr	38edd50c0e	Update copyright end year to 2017 in preparation for 3.13 release. n-i-bz git-svn-id: svn://svn.valgrind.org/valgrind/trunk@16333	2017-05-04 15:09:39 +00:00
Florian Krohm	193f88fad4	Make sure no executable stack gets created. Explanation by Matthias Schwarzott: The linker will request an executable stack as soon as at least one object file, that is linked in, wants an executable stack. And the absence of the .section .note.GNU-stack."",@progbits is enough to tell the linker that an executable stack is needed. So even an empty asm-file must at least contain this statement to not force executable stacks on the whole executable. * Define a helper macro MARK_STACK_NO_EXEC that disables the executable stack. * Instantiate this macro unconditionally at the end of each asm file. Patch by Matthias Schwarzott <zzam@gentoo.org>. git-svn-id: svn://svn.valgrind.org/valgrind/trunk@15692	2015-09-30 20:30:48 +00:00
Julian Seward	adc2dafee9	Update copyright dates, to include 2015. No functional change. git-svn-id: svn://svn.valgrind.org/valgrind/trunk@15577	2015-08-21 11:32:26 +00:00
Julian Seward	dbf9b63605	Update copyright dates (20XY-2012 ==> 20XY-2013) git-svn-id: svn://svn.valgrind.org/valgrind/trunk@13658	2013-10-18 14:27:36 +00:00
Julian Seward	4a3633e266	Update copyright dates to include 2012. git-svn-id: svn://svn.valgrind.org/valgrind/trunk@12843	2012-08-05 15:46:46 +00:00
Julian Seward	4deeeb4aa6	Use 32-bit XIndir counter incs, instead of 64-bit, as per r12527. git-svn-id: svn://svn.valgrind.org/valgrind/trunk@12528	2012-04-21 23:12:07 +00:00
Julian Seward	6d68ec0346	Add translation chaining support for ppc32 (tested) and to a large extent for ppc64 (incomplete, untested) (Valgrind side) git-svn-id: svn://svn.valgrind.org/valgrind/branches/TCHAIN@12512	2012-04-20 00:14:02 +00:00
Julian Seward	712ee2547b	Make the return type of VG_(disp_run_translations) be void, rather than the HWord it was claimed to be. Inconsistency spotted by Philippe Waroquiers. git-svn-id: svn://svn.valgrind.org/valgrind/branches/TCHAIN@12486	2012-04-04 12:23:23 +00:00
Julian Seward	8b6f93641c	Add translation chaining support for amd64, x86 and ARM (Valgrind side). See #296422. git-svn-id: svn://svn.valgrind.org/valgrind/branches/TCHAIN@12484	2012-04-02 21:56:03 +00:00
Julian Seward	c96096ab24	Update all copyright dates, from 20xy-2010 to 20xy-2011. git-svn-id: svn://svn.valgrind.org/valgrind/trunk@12206	2011-10-23 07:32:08 +00:00
Julian Seward	f8e9ab3607	Remove another memory reference from the arm dispatcher loop, by using the fact that all {VG,VEX}_TRC_VALUES have their lowest bit set. All other targets can benefit from this trick too. git-svn-id: svn://svn.valgrind.org/valgrind/trunk@11781	2011-05-28 11:05:44 +00:00
Julian Seward	3bcd288100	Get rid of a bunch of loads in the arm dispatcher inner loops, and make some attempt to schedule for Cortex-A8. Improves overall IPC for none running perf/bz2.c "-O" from 0.879 to 0.925. git-svn-id: svn://svn.valgrind.org/valgrind/trunk@11780	2011-05-28 10:16:58 +00:00
Julian Seward	5e6d4577de	Change the TT_FAST hash function for from "insn_address >> 2" to "insn_address >> 1". The former is appropriate for ARM code, where all insns are 4-sized and 4-aligned, but not for Thumb code, where the minimum size and alignment is 2. The old scheme happened to work for Thumb (indeed, any hash function would), but caused huge amounts of conflict misses in the fast cache for some programs. The change has been observed to reduce conflict misses by up to 100 times, and in some cases, improves performance significantly for Thumb code. Performance of ARM code is unchanged or possibly a bit worse. git-svn-id: svn://svn.valgrind.org/valgrind/trunk@11716	2011-04-28 14:58:15 +00:00
Julian Seward	de04801515	Merge from branches/THUMB: rack renaming of guest_R15 to guest_R15T. Also, add extra FPSCR masking for FPSCR invariant state sanity checks. git-svn-id: svn://svn.valgrind.org/valgrind/trunk@11279	2010-08-22 12:03:45 +00:00
Julian Seward	9b0574dff8	Update copyright dates to 2010. git-svn-id: svn://svn.valgrind.org/valgrind/trunk@11121	2010-05-03 21:37:12 +00:00
Julian Seward	e9de458500	Merge from branches/ARM, all parts of the ARM-Linux port except for the changes to do with reading and using ELF and DWARF3 info. This breaks all targets except amd64-linux and x86-linux. git-svn-id: svn://svn.valgrind.org/valgrind/trunk@10982	2010-01-01 11:59:33 +00:00

18 Commits