ansaurus

Question

Answer 1

+6 A:

Anything written regarding boxing/unboxing prior to .NET generics can be taken with something of a grain of salt. Generic collection types have removed the need for boxing and unboxing of value types, which makes using structs in these situations more valuable.

As for your specific slowdown - we'd probably need to see some code.

Erik Forbes 2009-02-28 01:12:22

Figured something like that, but haven't arrays of a certain type always been "generic"? Or was in int[] an object[] internally in .NET 1.0? As for source code: I can hardly post the entire source here, but I'll see if I can dig up something relevant.

JulianR 2009-02-28 01:17:03

Yes, arrays were always (sort of) generic.

configurator 2009-02-28 16:43:17

Arrays are "strongly typed".

Andrew Hare 2009-03-01 00:57:21

Answer 2

+3 A:

Basically, don't make them too big, and pass them around by ref when you can. I discovered this the exact same way... By changing my Vector and Ray classes to structs.

With more memory being passed around, it's bound to cause cache thrashing.

TraumaPony 2009-02-28 01:12:51

I thought value types were boxed when you pass them by ref?

Erik Forbes 2009-02-28 01:13:47

No, not at all. It just passes a pointer, I believe. For small methods, e.g. addition and simple ray-interesection tests, the cost of pointer dereferencing is less than the cost of copying the memory.

TraumaPony 2009-02-28 01:17:12

Interesting, passing by ref would kill all my pretty operator overloading and elegant code. I'll try to benchmark this.

JulianR 2009-02-28 01:19:19

Passing by reference? The point of having a struct is that it's so small that it's more efficient to copy it than to have the overhead of a reference to it...

Guffa 2009-02-28 04:12:39

From my experience passing them as ref was slowing things down.

Joan Venge 2009-02-28 06:02:55

Necromancing this as I've grown a little more experienced since I asked this. In a rewrite of the raytracer, passing *all* structs by reference was indeed a lot faster. But yeah, it really did make my code very ugly. Small mathematical one-lined statements turned into 5 lined blocks of Vector3.Add/Subtract/Multiply/etc and I had to resort to a lot of public fields because a property can't be passed by ref (makes sense, it's a method after all). It's really fast now though, real-time framerates in very simple scenes.

JulianR 2009-06-29 00:50:36

Answer 3

+5 A:

I think the key lies in these two statements from your post:

you create millions of them

and

I do pass them to methods when needed of course

Now unless your struct is less than or equal to 4 bytes in size (or 8 bytes if you are on a 64-bit system) you are copying much more on each method call then if you simply passed an object reference.

Andrew Hare 2009-02-28 01:17:54

It's still quicker than the massive amount of garbage collection that would occur...

TraumaPony 2009-02-28 01:18:54

Apparently not :)

Andrew Hare 2009-02-28 01:21:36

Does it actually do any GC though? If I render a really large image, the memory usage of my process just keeps climbing indefinitely, perhaps because even at using 3 GB, I still have plenty of memory left and perhaps in that case the GC rather waits until I'm done hogging the CPU.

JulianR 2009-02-28 01:29:40

Well, it should be GCing. At least, that was the problem with my raytracer.

TraumaPony 2009-02-28 01:33:48

@Andrew Hare: Nice

Ed Swangren 2009-02-28 02:08:19

I agree with TraumaPony. I had the same problem with a 3d renderer I wrote. When things are used and destroyed in an instant, structs really make things way faster. Also if you check out XNA for instance almost everything I have used was a struct.

Joan Venge 2009-02-28 06:00:11

That is only true if you are not passing those structs to other methods.

Andrew Hare 2009-02-28 12:43:24

Try passing these structs by ref and see what happens.

NM 2009-02-28 17:23:46

Profile the app with a .NET Memory Profiler

Andrei Rinea 2009-03-03 23:43:13

Answer 4

A:

You can also make structs into Nullable objects. Custom classes will not be able to created

as

Nullable<MyCustomClass> xxx = new Nullable<MyCustomClass>

where with a struct is nullable

Nullable<MyCustomStruct> xxx = new Nullable<MyCustomStruct>

But you will be (obviously) losing all your inheritance features

BozoJoe 2009-02-28 01:26:43

This just wraps the struct in a class, thus invoking the GC. You're better off just changing your struct to a class, as it would have the same effect and is probably slightly less confusing

Grant Peters 2009-02-28 16:18:47

Answer 5

+1 A:

Have you profiled the application? Profiling is the only sure fire way to see where the actual performance problem is. There are operations that are generally better/worse on structs but unless you profile you'd just be guessing as to what the problem is.

JaredPar 2009-02-28 01:40:30

Answer 6

+4 A:

The first thing I would look for is to make sure that you have explicitly implemented Equals and GetHashCode. Failing to do this means that the runtime implementation of each of these does some very expensive operations to compare two struct instances (internally it uses reflection to determine each of the private fields and then checkes them for equality, this causes a significant amount of allocation).

Generally though, the best thing you can do is to run your code under a profiler and see where the slow parts are. It can be an eye-opening experience.

2009-02-28 03:12:03

I tried the overrides, even though I didn't use any Vector or Ray comparisons, but it had no effect. Good tip though, I will make sure to override Equals and GetHashCode from now on :)

JulianR 2009-02-28 12:25:06

Answer 7

+3 A:

In the recommendations for when to use a struct it says that it should not be larger than 16 bytes. Your Vector is 12 bytes, which is close to the limit. The Ray has two Vectors, putting it at 24 bytes, which is clearly over the recommended limit.

When a struct gets larger than 16 bytes it can no longer be copied efficiently with a single set of instructions, instead a loop is used. So, by passing this "magic" limit, you are actually doing a lot more work when you pass a struct than when you pass a reference to an object. This is why the code is faster with classes eventhough there is more overhead when allocating the objects.

The Vector could still be a struct, but the Ray is simply too large to work well as a struct.

Guffa 2009-02-28 04:04:24

I see, didn't know the 16 bytes limit was such a hard limit, thought of it more as a guideline. However, the funny thing is that having my Ray as struct doesn't slow the application down much if at all, while making my Vector a struct does slow it down by 50%.

JulianR 2009-02-28 12:16:39

Having Vector as a class and Ray as a struct will make Ray contain two references. That will work sizewise, but you may get some surprising semantic effects. Making both structs is what puts the Ray struct over the size limit.

Guffa 2009-02-28 14:29:52

Answer 8

+2 A:

See also

http://stackoverflow.com/questions/521298/when-to-use-struct-in-c/521343

(especially my answer there :) )

Brian 2009-02-28 05:49:13

Answer 9

A:

I use structs basically for parameter objects, returning multiple pieces of information from a function, and... nothing else. Don't know whether it's "right" or "wrong," but that's what I do.

Daniel Straight 2009-02-28 06:05:38

Answer 10

+4 A:

An array of structs would be a single contiguous structure in memory, while items in an array of objects (instances of reference types) need to be individually addressed by a pointer (i.e. a reference to an object on the garbage-collected heap). Therefore if you address large collections of items at once, structs will give you a performance gain since they need fewer indirections. In addition, structs cannot be inherited, which might allow the compiler to make additional optimizations (but that is just a possibility and depends on the compiler).

However, structs have quite different assignment semantics and also cannot be inherited. Therefore I would usually avoid structs except for the given performance reasons when needed.

struct

An array of values v encoded by a struct (value type) looks like this in memory:

vvvv

class

An array of values v encoded by a class (reference type) look like this:

pppp

..v..v...v.v..

where p are the this pointers, or references, which point to the actual values v on the heap. The dots indicate other objects that may be interspersed on the heap. In the case of reference types you need to reference v via the corresponding p, in the case of value types you can get the value directly via its offset in the array.

ILoveFortran 2009-02-28 15:46:17

Answer 11

A:

If the structs are small, and not too many exist at once, it SHOULD be placing them on the stack (as long as its a local variable and not a member of a class) and not on the heap, this means the GC doesn't need to be invoked and memory allocation/deallocation should be almost instantaneous.

When passing a struct as a parameter to function, the struct is copied, which not only means more allocations/deallocations (from the stack, which is almost instantaneous, but still has overhead), but the overhead in just transferring data between the 2 copies. If you pass via reference, this is a non issue as you are just telling it where to read the data from, rather than copying it.

I'm not 100% sure on this, but i suspect that returning arrays via an 'out' parameter may also give you a speed boost, as memory on the stack is reserved for it and doesn't need to be copied as the stack is "unwound" at the end of function calls.

Grant Peters 2009-02-28 16:14:15

I often hear that structs live on the stack, not on the garbage collected heap; that is not quite correct. They don't live alone but they can certainly live on the heap as part of another reference-type object.

ILoveFortran 2009-02-28 16:36:00

When passing a struct via copy no further allocation is needed - the corresponding amount is reserved on the stack before the call. The copying still needs to happen, though, that's what makes passing structs more expensive, not the allocation.

ILoveFortran 2009-02-28 16:42:42

I did say that it SHOULD be placed on there under the assumption that they are local variable (as the was detailed in the question).Also, memory does need to actually be allocated from the stack as well, its basically a non issue, but there is still slight overhead.

Grant Peters 2009-03-01 06:18:04

Answer 12

A:

My own ray tracer also uses struct Vectors (though not Rays) and changing Vector to class does not appear to have any impact on the performance. I'm currently using three doubles for the vector so it might be bigger than it ought to be. One thing to note though, and this might be obvious but it wasn't for me, and that is to run the program outside of visual studio. Even if you set it to optimized release build you can get a massive speed boost if you start the exe outside of VS. Any benchmarking you do should take this into consideration.

Morten Christiansen 2009-04-17 10:13:03

Answer 13

A:

While the functionality is similar, structures are usually more efficient than classes. You should define a structure, rather than a class, if the type will perform better as a value type than a reference type.

Specifically, structure types should meet all of these criteria:

Logically represents a single value
Has an instance size less than 16 bytes
Will not be changed after creation
Will not be cast to a reference type

Gineer 2009-04-17 15:00:26

ansaurus

tags:

views:

answers:

When are structs the answer?

related questions